17 citations · 17 across the 3 of their papers we have counts for
1 paper · 1 filter
OpenAI, :, Ahmed El-Kishky +23
We show that reinforcement learning applied to large language models (LLMs) significantly boosts performance on complex coding and reasoning tasks. Additionally, we compare two gen…