2 papers
cs.LG2026
Position: RL Researchers Need to Distinguish Between Solving Simulators and Using Simulators as a Proxy
Matthew Vandergrift, Esraa Elelimy, Martha White
One goal in reinforcement learning (RL) research is to understand general-purpose sequential decision-making, using benchmark simulators as a proxy for learning in deployment setti…
cs.LG2026
Measure-to-measure Regression with Transformers
Matthew Vandergrift, Martha White, Yury Polyanskiy +2
Many learning problems require predicting how populations evolve under an unknown transformation. A natural representation for such populations is a probability measure, with point…