3 papers
cs.AI2026
Trust but Verify: Prover-Verifier Deliberation for Selective LLM Prediction
João Sedoc, Baotong Zhang, Dean Foster
Reliably knowing when a language model is correct is almost as important as being correct. We introduce prover-verifier deliberation (PVD), an inference-time protocol grounded in i…
cs.LG2026
Optimal Budgeted Adaptation of Large Language Models
Jing Wang, Jie Shen, Dean Foster +2
The trade-off between labeled data availability and downstream accuracy remains a central challenge in fine-tuning large language models (LLMs). We propose a principled framework f…
cs.GT2023
Playing Large Games with Oracles and AI Debate
Xinyi Chen, Angelica Chen, Dean Foster +1
We consider regret minimization in repeated games with a very large number of actions. Such games are inherent in the setting of AI Safety via Debate \cite{irving2018ai}, and more…