2 papers
cs.CL2026
Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning
Yu Xia, Zhouhang Xie, Xin Xu +4
Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little inference-time control. Exist…
cs.CL2025
Improving In-Context Learning with Reasoning Distillation
Nafis Sadeq, Xin Xu, Zhouhang Xie +4
Language models rely on semantic priors to perform in-context learning, which leads to poor performance on tasks involving inductive reasoning. Instruction-tuning methods based on…