1 paper
Keuntae Kim, Eunhye Jeong, Sehyeon Lee +2
Recent advances in enhancing the reasoning ability of large language models (LLMs) have been remarkably successful. LLMs trained with reinforcement learning (RL) for reasoning demo…