4.3k citations · 4.4k across the 5 of their papers we have counts for
1 paper · 2 filters
OpenAI, :, Aaron Jaech +261
The o1 model series is trained with large-scale reinforcement learning to reason using chain of thought. These advanced reasoning capabilities provide new avenues for improving the…