Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
Zheng Li, Qingxiu Dong, Jingyuan Ma +3
Recently, large reasoning models demonstrate exceptional performance on various tasks. However, reasoning models always consume excessive tokens even for simple queries, leading to…
cs.AI2026
Decoding in Geometry: Alleviating Embedding-Space Crowding for Complex Reasoning
Yixin Yang, Qingxiu Dong, Zhifang Sui
Sampling-based decoding underlies complex reasoning in large language models (LLMs), where decoding strategies critically shape model behavior. Temperature- and truncation-based me…