1 paper
Jinyoung Park, Jeehye Na, Jinyoung Kim +1
Recent works have demonstrated the effectiveness of reinforcement learning (RL)-based post-training for enhancing the reasoning capabilities of large language models (LLMs). In par…