1 paper · 1 filter
Jie Cao, Tianwei Lin, Zhenxuan Fan +5
Long chain-of-thought~(CoT) has become a dominant paradigm for enhancing the reasoning capability of large reasoning models~(LRMs); however, the performance gains often come with a…