7 citations · 7 across the 1 of their papers we have counts for
1 paper · 1 filter
Jinghang Han, Jiawei Chen, Hang Shao +7
Reinforcement learning has significantly enhanced the reasoning capabilities of Large Language Models (LLMs) in complex problem-solving tasks. Recently, the introduction of DeepSee…