1 paper · 1 filter
Valentin Lacombe, Valentin Quesnel, Damien Sileo
We introduce Reasoning Core, a new scalable environment for Reinforcement Learning with Verifiable Rewards (RLVR), designed to advance foundational symbolic reasoning in Large Lang…