1 citations · 1 across the 2 of their papers we have counts for
1 paper · 1 filter
Rui Sun, Yifan Sun, Sheng Xu +5
Reinforcement Learning (RL) has enabled Large Language Models (LLMs) to achieve remarkable reasoning in domains like mathematics and coding, where verifiable rewards provide clear…