1 paper · 1 filter
Tian Qin, Core Francisco Park, Mujin Kwun +5
Mathematical reasoning tasks have become prominent benchmarks for assessing the reasoning capabilities of LLMs, especially with reinforcement learning (RL) methods such as GRPO sho…