Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
LongR: Unleashing Long-Context Reasoning via Reinforcement Learning with Dense Utility Rewards
Bowen Ping, Zijun Chen, Yiyao Yu +3
Reinforcement Learning has emerged as a key driver for LLM reasoning. This capability is equally pivotal in long-context scenarios--such as long-dialogue understanding and structur…
cs.CL2025
ProtoReasoning: Prototypes as the Foundation for Generalizable Reasoning in LLMs
Feng He, Zijun Chen, Xinnian Liang +4
Recent advances in Large Reasoning Models (LRMs) trained with Long Chain-of-Thought (Long CoT) reasoning have demonstrated remarkable cross-domain generalization capabilities. Howe…