4 papers · 1 filter
Beyond Fast and Slow: Cognitive-Inspired Elastic Reasoning for Large Language Models
Jinwu Hu, Dongjin Yang, Langyu Bian +6
Large language models (LLMs) have demonstrated impressive performance across various language tasks. However, existing LLM reasoning strategies mainly rely on the LLM itself with f…
Continual Knowledge Adaptation for Reinforcement Learning
Jinwu Hu, Zihao Lian, Zhiquan Wen +5
Reinforcement Learning enables agents to learn optimal behaviors through interactions with environments. However, real-world environments are typically non-stationary, requiring ag…
Beyond Model Scaling: Test-Time Intervention for Efficient Deep Reasoning
Qianyue Wang, Jinwu Hu, Yufeng Wang +5
Large Reasoning Models (LRMs) excel at multi-step reasoning but often suffer from inefficient reasoning processes like overthinking and overshoot, where excessive or misdirected re…
Enhancing User-Oriented Proactivity in Open-Domain Dialogues with Critic Guidance
Yufeng Wang, Jinwu Hu, Ziteng Huang +10
Open-domain dialogue systems aim to generate natural and engaging conversations, providing significant practical value in real applications such as social robotics and personal ass…