1 paper · 1 filter
Chongyu Fan, Yihua Zhang, Jinghan Jia +2
Large reasoning models (LRMs), such as OpenAI's o1 and DeepSeek-R1, harness test-time scaling to perform multi-step reasoning for complex problem-solving. This reasoning process, e…