1 paper · 1 filter
Zihan Wang, Cheng Tang, Lei Gong +5
Chain-of-Thought (CoT) reasoning in large language models (LLMs) significantly improves accuracy on complex tasks, yet incurs excessive memory overhead due to the long think-stage…