1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Chuou Xu, Liya Ji, Qifeng Chen
Reinforcement learning (RL) as post-training is crucial for enhancing the reasoning ability of large language models (LLMs) in coding and math. However, their capacity for visual s…