17 citations · 37 across the 10 of their papers we have counts for
Showing cs.GTShow all
2 papers · 1 filter
cs.GT2025
Fundamental Limits of Game-Theoretic LLM Alignment: Smith Consistency and Preference Matching
Zhekun Shi, Kaizhao Liu, Qi Long +2
Nash Learning from Human Feedback is a game-theoretic framework for aligning large language models (LLMs) with human preferences by modeling learning as a two-player zero-sum game.…
cs.GT2025
Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium
Kaizhao Liu, Qi Long, Zhekun Shi +2
Aligning large language models (LLMs) with diverse human preferences is critical for ensuring fairness and informed outcomes when deploying these models for decision-making. In thi…