2 papers
cs.CL2026
A*-Thought-V2: Efficient Latent Reasoning via Geometric Dynamics of LLM
Xiaoang Xu, Siyuan Liu, Shuo Wang +13
Chain-of-Thought (CoT) improves the reasoning ability of Large Language Models (LLMs) but incurs substantial computation and context costs. Existing methods either lose intermediat…
cs.MA2026
DREvo: Distilling Recalibrated Historical Experience for Harness Self-Evolution
Hanghui Guo, Weijie Shi, Zhangze Chen +6
Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Therefore, recent research has i…