3 papers
cs.AI2025
Synthetic Error Injection Fails to Elicit Self-Correction In Language Models
David X. Wu, Shreyas Kapur, Anant Sahai +1
Reinforcement learning has become the dominant paradigm for eliciting reasoning and self-correction capabilities in large language models, but its computational expense motivates e…
cs.LG2025
Precise Asymptotic Generalization for Multiclass Classification with Overparameterized Linear Models
David X. Wu, Anant Sahai
We study the asymptotic generalization of an overparameterized linear model for multiclass classification under the Gaussian covariates bi-level model introduced in Subramanian et…
cs.LG2025
Provable Weak-to-Strong Generalization via Benign Overfitting
David X. Wu, Anant Sahai
The classic teacher-student model in machine learning posits that a strong teacher supervises a weak student to improve the student's capabilities. We instead consider the inverted…