Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
COMPOSITE-Stem
Kyle Waters, Lucas Nuzzi, Tadhg Looram +20
AI agents hold growing promise for accelerating scientific discovery; yet, a lack of frontier evaluations hinders adoption into real workflows. Expert-written benchmarks have prove…
cs.AI2025
Reasoning Core: A Scalable RL Environment for LLM Symbolic Reasoning
Valentin Lacombe, Valentin Quesnel, Damien Sileo
We introduce Reasoning Core, a new scalable environment for Reinforcement Learning with Verifiable Rewards (RLVR), designed to advance foundational symbolic reasoning in Large Lang…