2 papers
cs.CL2026
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
Gao Yang, Yuhang Liu, Siyu Miao +3
Ideal or real - that is the question.In this work, we explore whether principles from game theory can be effectively applied to the evaluation of large language models (LLMs). This…
cs.IR2024
Subtopic-aware View Sampling and Temporal Aggregation for Long-form Document Matching
Youchao Zhou, Heyan Huang, Zhijing Wu +2
Long-form document matching aims to judge the relevance between two documents and has been applied to various scenarios. Most existing works utilize hierarchical or long context mo…