Showing cs.LGShow all
3 papers · 1 filter
cs.LG2026
Progressive Agent Skill Generation via Reinforcement Learning
Junhao Shen, Zhanqiu Zhang, Yiwen Guo +1
Recent large language model agents often use external skills as modular procedural units that condition inference and improve complex task solving. Thus, automatically generating h…
cs.LG2026
Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning
Junhao Shen, Teng Zhang, Xiaoyan Zhao +1
Large language model agents increasingly rely on external skills to solve complex tasks, where skills act as modular units that extend their capabilities beyond what parametric mem…
cs.LG2025
Semi-off-Policy Reinforcement Learning for Vision-Language Slow-Thinking Reasoning
Junhao Shen, Haiteng Zhao, Yuzhe Gu +7
Enhancing large vision-language models (LVLMs) with visual slow-thinking reasoning is crucial for solving complex multimodal tasks. However, since LVLMs are mainly trained with vis…