2 papers
cs.CL2026
Inference-time Alignment in Continuous Space
Yige Yuan, Teng Xiao, Li Yunfan +5
Aligning large language models with human feedback at inference time has received increasing attention due to its flexibility. Existing methods rely on generating multiple response…
cs.CL2025
ProtoReasoning: Prototypes as the Foundation for Generalizable Reasoning in LLMs
Feng He, Zijun Chen, Xinnian Liang +4
Recent advances in Large Reasoning Models (LRMs) trained with Long Chain-of-Thought (Long CoT) reasoning have demonstrated remarkable cross-domain generalization capabilities. Howe…