2 papers
stat.ML2026
A Judge-Aware Ranking Framework for Evaluating Large Language Models without Ground Truth
Mingyuan Xu, Xinzi Tan, Jiawei Wu +1
Evaluating large language models (LLMs) on open-ended tasks without ground-truth labels is increasingly done via the LLM-as-a-judge paradigm. A critical but under-modeled issue is…
cs.LG2026
From Hawkes Processes to Attention: Time-Modulated Mechanisms for Event Sequences
Xinzi Tan, Kejian Zhang, Junhan Yu +1
Marked Temporal Point Processes (MTPPs) arise naturally in medical, social, commercial, and financial domains. However, existing Transformer-based methods mostly inject temporal in…