3 papers
cs.LG2025
What is the objective of reasoning with reinforcement learning?
Damek Davis, Benjamin Recht
We show that several popular algorithms for reinforcement learning in large language models with binary rewards can be viewed as stochastic gradient ascent on a monotone transform…
stat.OT2025
The Actuary's Final Word on Algorithmic Decision Making
Benjamin Recht
Paul Meehl's foundational work "Clinical versus Statistical Prediction," provided early theoretical justification and empirical evidence of the superiority of statistical methods o…
eess.SY2025
On Sampling Time and Invariance
Spencer Schutz, Charlott Vallon, Ben Recht +1
Invariant sets define regions of the state space where system constraints are always satisfied. The majority of numerical techniques for computing invariant sets have been develope…