2 papers
cs.LG2026
-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data
Harsh Goel, Akhil Udathu, Susmija Jabbireddy +2
Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle primarily due to limitations…
cs.LG2025
A Personalized Exercise Assistant using Reinforcement Learning (PEARL): Results from a four-arm Randomized-controlled Trial
Amy Armento Lee, Narayan Hegde, Nina Deliu +16
Consistent physical inactivity poses a major global health challenge. Mobile health (mHealth) interventions, particularly Just-in-Time Adaptive Interventions (JITAIs), offer a prom…