198 citations · 220 across the 29 of their papers we have counts for
Showing cs.AIShow all
3 papers · 1 filter
cs.AI2026
Offline-Online Curriculum RL for Multimodal Reasoning
Wendi Deng, Hang Du, Guoshun Nan +11
Multimodal large language models exhibit capabilities on reasoning tasks, yet often produce flawed intermediate steps while yielding correct final answers. This behavior undermines…
cs.AI2026
AutoRAS: Learning Robust Agentic Systems with Primitive Representations
Yang Yue, Xuancheng Zhu, Yuyang Ma +7
The automated design of agentic systems offers a promising pathway for scaling large language models (LLMs) beyond single-agent reasoning. While prior work has advanced task perfor…
cs.AI2026
Can LLM Agents Sustain Long-Horizon Organizational Dynamics?
Xuancheng Zhu, Yang Yue, Shuaibing Wan +4
Large language agents are increasingly used for social simulation, yet it remains unclear whether they can sustain coherent behavior in structured organizations, where goals must p…