1 paper · 1 filter
Yu Xia, Zhouhang Xie, Xin Xu +4
Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little inference-time control. Exist…