2 papers
cs.AI2025
Reasoning Core: A Scalable RL Environment for LLM Symbolic Reasoning
Valentin Lacombe, Valentin Quesnel, Damien Sileo
We introduce Reasoning Core, a new scalable environment for Reinforcement Learning with Verifiable Rewards (RLVR), designed to advance foundational symbolic reasoning in Large Lang…
cs.CL2025
Saturation-Driven Dataset Generation for LLM Mathematical Reasoning in the TPTP Ecosystem
Valentin Quesnel, Damien Sileo
The scarcity of high-quality, logically sound data is a critical bottleneck for advancing the mathematical reasoning of Large Language Models (LLMs). Our work confronts this challe…