1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025
Local Coherence or Global Validity? Investigating RLVR Traces in Math Domains
Soumya Rani Samineni, Durgesh Kalwar, Vardaan Gangal +2
Reinforcement Learning with Verifiable Rewards (RLVR)-based post-training of Large Language Models (LLMs) has been shown to improve accuracy on reasoning tasks and continues to att…
cs.LG2023★ 1 cited
CoRL: Environment Creation and Management Focused on System Integration
Justin D. Merrick, Benjamin K. Heiner, Cameron Long +5
Existing reinforcement learning environment libraries use monolithic environment classes, provide shallow methods for altering agent observation and action spaces, and/or are tied…