4 papers
CoSkill: Joint Reinforcement Learning of Reasoning and Meta-Skill Agents for Hierarchical Skill Evolution
Jinyuan Feng, Dongmin Li, Yiqun Chen +4
Skill libraries improve the sample efficiency of agentic reinforcement learning (RL) by enabling large language model (LLM) agents to reuse procedural knowledge. Yet existing parad…
Focus on the Core: Empowering Diffusion Large Language Models by Self-Contrast
Jinyuan Feng, Xin Yu, Yiqun Chen +5
The iterative denoising paradigm of Diffusion Large Language Models (DLMs) endows them with a distinct advantage in global context modeling. However, current decoding strategies fa…
Self-Compression of Chain-of-Thought via Multi-Agent Reinforcement Learning
Yiqun Chen, Jinyuan Feng, Wei Yang +9
The inference overhead induced by redundant reasoning undermines the interactive experience and severely bottlenecks the deployment of Large Reasoning Models. Existing reinforcemen…
TacEleven: generative tactic discovery for football open play
Siyao Zhao, Hao Ma, Zhiqiang Pu +4
Creating offensive advantages during open play is fundamental to football success. However, due to the highly dynamic and long-sequence nature of open play, the potential tactic sp…