2 papers
cs.CL2026
Self-Reflective Generation at Test Time
Jian Mu, Qixin Zhang, Zhiyong Wang +5
Large language models (LLMs) increasingly solve complex reasoning tasks via long chain-of-thought, but their forward-only autoregressive generation process is fragile; early token…
cs.CL2025
Thinking with Nothinking Calibration: A New In-Context Learning Paradigm in Reasoning Large Language Models
Haotian Wu, Bo Xu, Yao Shu +2
Reasoning large language models (RLLMs) have recently demonstrated remarkable capabilities through structured and multi-step reasoning. While prior research has primarily focused o…