5 papers
Adaptive Coarse-to-Fine Subgoal Refinement for Long-Horizon Offline Goal-Conditioned Reinforcement Learning
Kaiqiang Ke, Shenghong He, Chengdong Xu +3
Offline goal-conditioned reinforcement learning (GCRL) is challenging in long-horizon tasks, where distant state--goal pairs provide weak supervision and value estimates become vul…
EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems
Chengdong Xu, Kaiqiang Ke, Ziheng Liu +4
Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborative reasoning. However, most…
Context-Picker: Dynamic context selection using multi-stage reinforcement learning
Siyuan Zhu, Chengdong Xu, Kaiqiang Ke +1
In long-context question answering, selecting the appropriate scope of context for a query remains a key and unresolved challenge. Insufficient context can lead to missing essentia…
HR: Hierarchical Hindsight Reflection for Multi-Task LLM Agents
Shicheng Ye, Chao Yu, Kaiqiang Ke +2
Large language model (LLM)-based agents have shown strong potential in multi-task scenarios, owing to their ability to transfer knowledge across diverse tasks. However, existing ap…
GCHR : Goal-Conditioned Hindsight Regularization for Sample-Efficient Reinforcement Learning
Xing Lei, Wenyan Yang, Kaiqiang Ke +4
Goal-conditioned reinforcement learning (GCRL) with sparse rewards remains a fundamental challenge in reinforcement learning. While hindsight experience replay (HER) has shown prom…