2 papers
cs.LG2025
Doubly-Robust LLM-as-a-Judge: Externally Valid Estimation with Imperfect Personas
Luke Guerdan, Justin Whitehouse, Kimberly Truong +2
As Generative AI (GenAI) systems see growing adoption, a key concern involves the external validity of evaluations, or the extent to which they generalize from lab-based to real-wo…
cs.LG2025
Nearly-Optimal Bandit Learning in Stackelberg Games with Side Information
Maria-Florina Balcan, Martino Bernasconi, Matteo Castiglioni +3
We study the problem of online learning in Stackelberg games with side information between a leader and a sequence of followers. In every round the leader observes contextual infor…