Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
COMAP: Co-Evolving World Models and Agent Policies for LLM Agents
Youwei Liu, Jian Wang, Hanlin Wang +1
Equipping language agents with world models enables them to anticipate environment dynamics and evaluate candidate actions before execution. However, existing textual world models…
cs.AI2026
Finding RELIEF: Shaping Reasoning Behavior without Reasoning Supervision via Belief Engineering
Chak Tou Leong, Dingwei Chen, Heming Xia +4
Large reasoning models (LRMs) have achieved remarkable success in complex problem-solving, yet they often suffer from computational redundancy or reasoning unfaithfulness. Current…
cs.AI2025
Scaling over Scaling: Exploring Test-Time Scaling Plateau in Large Reasoning Models
Jian Wang, Boyan Zhu, Chak Tou Leong +2
Large reasoning models (LRMs) have exhibited the capacity of enhancing reasoning performance via internal test-time scaling. Building upon this, a promising direction is to further…