1 citations · 1 across the 6 of their papers we have counts for
1 paper · 1 filter
Ziru Chen, Dongdong Chen, Ruinan Jin +3
Recently, there have been significant research interests in training large language models (LLMs) with reinforcement learning (RL) on real-world tasks, such as multi-turn code gene…