1 citations · 2 across the 9 of their papers we have counts for
1 paper · 2 filters
Licheng Liu, Zihan Wang, Linjie Li +5
Multi-turn problem solving is critical yet challenging for Large Reasoning Models (LRMs) to reflect on their reasoning and revise from feedback. Existing Reinforcement Learning (RL…