1 paper
Kaiwen Wei, Rui Shan, Dongsheng Zou +4
Large reasoning models (LRMs) have shown significant progress in test-time scaling through chain-of-thought prompting. Current approaches like search-o1 integrate retrieval augment…