5 papers
CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning
Junru Song, Wenhao Zhang, Yang Yang +7
Target-oriented vision-and-language navigation (VLN) on aerial platforms is attracting growing attention for missions such as disaster rescue, infrastructure inspection, and securi…
HiMAC: Hierarchical Macro-Micro Learning for Long-Horizon LLM Agents
Hongbo Jin, Rongpeng Zhu, Jiayu Ding +2
Large language model (LLM) agents have recently demonstrated strong capabilities in interactive decision-making, yet they remain fundamentally limited in long-horizon tasks that re…
Towards Monotonic Improvement in In-Context Reinforcement Learning
Wenhao Zhang, Shao Zhang, Xihuai Wang +2
In-Context Reinforcement Learning (ICRL) has emerged as a promising paradigm for developing agents that can rapidly adapt to new tasks by leveraging past experiences as context, wi…
Leveraging Dual Process Theory in Language Agent Framework for Real-time Simultaneous Human-AI Collaboration
Shao Zhang, Xihuai Wang, Wenhao Zhang +10
Agents built on large language models (LLMs) have excelled in turn-by-turn human-AI collaboration but struggle with simultaneous tasks requiring real-time interaction. Latency issu…
Aligning Individual and Collective Objectives in Multi-Agent Cooperation
Yang Li, Wenhao Zhang, Jianhong Wang +4
Among the research topics in multi-agent learning, mixed-motive cooperation is one of the most prominent challenges, primarily due to the mismatch between individual and collective…