6 papers
MobileGym: A Verifiable and Highly Parallel Simulation Platform for Mobile GUI Agent Research
Dingbang Wu, Rui Hao, Haiyang Wang +8
We present MobileGym, a browser-hosted, lightweight, fully controllable environment for everyday mobile use, targeting interaction fidelity without replicating proprietary backends…
Reinforced Reasoning for Embodied Planning
Di Wu, Jiaxin Fan, Junzhe Zang +4
Embodied planning requires agents to make coherent multi-step decisions based on dynamic visual observations and natural language goals. While recent vision-language models (VLMs)…
TextAtari: 100K Frames Game Playing with Language Agents
Wenhao Li, Wenwu Li, Chuyun Shen +8
We present TextAtari, a benchmark for evaluating language agents on very long-horizon decision-making tasks spanning up to 100,000 steps. By translating the visual state representa…
SentinelAgent: Graph-based Anomaly Detection in Multi-Agent Systems
Xu He, Di Wu, Yan Zhai +1
The rise of large language model (LLM)-based multi-agent systems (MAS) introduces new security and reliability challenges. While these systems show great promise in decomposing and…
Exploring the Necessity of Reasoning in LLM-based Agent Scenarios
Xueyang Zhou, Guiyao Tie, Guowen Zhang +7
The rise of Large Reasoning Models (LRMs) signifies a paradigm shift toward advanced computational reasoning. Yet, this progress disrupts traditional agent frameworks, traditionall…
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
Di Wu, Xian Wei, Guang Chen +4
Embodied multi-agent systems (EMAS) have attracted growing attention for their potential to address complex, real-world challenges in areas such as logistics and robotics. Recent a…