2 papers
cs.LG2025
Generative Classifiers Avoid Shortcut Solutions
Alexander C. Li, Ananya Kumar, Deepak Pathak
Discriminative approaches to classification often learn shortcuts that hold in-distribution but fail even under minor distribution shift. This failure mode stems from an overrelian…
cs.CL2025
Language Models Prefer What They Know: Relative Confidence Estimation via Confidence Preferences
Vaishnavi Shrivastava, Ananya Kumar, Percy Liang
Language models (LMs) should provide reliable confidence estimates to help users detect mistakes in their outputs and defer to human experts when necessary. Asking a language model…