2 papers
cs.CL2026
LLMs Judge Themselves: A Game-Theoretic Framework for Human-Aligned Evaluation
Gao Yang, Yuhang Liu, Siyu Miao +3
Ideal or real - that is the question.In this work, we explore whether principles from game theory can be effectively applied to the evaluation of large language models (LLMs). This…
cs.LG2024
Online Sequential Decision-Making with Unknown Delays
Ping Wu, Heyan Huang, Zhengyang Liu
In the field of online sequential decision-making, we address the problem with delays utilizing the framework of online convex optimization (OCO), where the feedback of a decision…