1 paper
Shu Zhou, Rui Ling, Junan Chen +3
Scaling test-time compute through extended chains of thought has become a dominant paradigm for improving large language model reasoning. However, existing research implicitly assu…