1 paper
Haoran Xu, Hongyu Wang, Jiaze Li +5
Existing LLM test-time scaling laws emphasize the emergence of self-reflective behaviors through extended reasoning length. Nevertheless, this vertical scaling strategy often encou…