10 citations · 10 across the 1 of their papers we have counts for
1 paper · 1 filter
Xiaoxi Li, Guanting Dong, Jiajie Jin +5
Large reasoning models (LRMs) like OpenAI-o1 have demonstrated impressive long stepwise reasoning capabilities through large-scale reinforcement learning. However, their extended r…