7 citations · 7 across the 3 of their papers we have counts for
1 paper · 1 filter
Wenzhe Niu, Wei He, Zongxia Xie +10
Reinforcement learning has become a cornerstone for enhancing the reasoning capabilities of Large Language Models, where group-based approaches such as GRPO have emerged as efficie…