1 paper
Xinhan Di, JoyJiaoW
Reinforcement learning scaling enhances the reasoning capabilities of large language models, with reinforcement learning serving as the key technique to draw out complex reasoning.…