Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
Self-Improving is Often Sudden: Enlightenment-style Finetuning for Large-Scale Models
Jing-Xiao Liao, Tianwei Zhang, Yu-Hao Jiang +3
The pursuit of autonomously self-improving models has attracted growing interest in the era of large-scale foundation models. Drawing inspiration from the concept of "enlightenment…
cs.LG2025
BudgetThinker: Empowering Budget-aware LLM Reasoning with Control Tokens
Hao Wen, Xinrui Wu, Yi Sun +7
Recent advancements in Large Language Models (LLMs) have leveraged increased test-time computation to enhance reasoning capabilities, a strategy that, while effective, incurs signi…