Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Calibeating Made Simple
Yurong Chen, Zhiyi Huang, Michael I. Jordan +1
We study calibeating, the problem of post-processing external forecasts online to minimize cumulative losses and match an informativeness-based benchmark. Unlike prior work, which…
cs.LG2026
How Sampling Shapes LLM Alignment: From One-Shot Optima to Iterative Dynamics
Yurong Chen, Yu He, Michael I. Jordan +1
Standard methods for aligning large language models with human preferences learn from pairwise comparisons among sampled candidate responses and regularize toward a reference polic…