1 citations · 2 across the 6 of their papers we have counts for
1 paper · 1 filter
Shiye Lei, Zhihao Cheng, Kai Jia +1
Large language models (LLMs) have recently demonstrated remarkable progress in reasoning capabilities through reinforcement learning with verifiable rewards (RLVR). By leveraging s…