4 papers
Hydra-Nav: Object Navigation via Adaptive Dual-Process Reasoning
Zixuan Wang, Huang Fang, Shaoan Wang +4
While large vision-language models (VLMs) show promise for object goal navigation, current methods still struggle with low success rates and inefficient localization of unseen obje…
ANNIE: Be Careful of Your Robots
Yiyang Huang, Zixuan Wang, Zishen Wan +4
The integration of vision-language-action (VLA) models into embodied AI (EAI) robots is rapidly advancing their ability to perform complex, long-horizon tasks in humancentric envir…
KARMA: Augmenting Embodied AI Agents with Long-and-short Term Memory Systems
Zixuan Wang, Bo Yu, Junzhe Zhao +6
Embodied AI agents responsible for executing interconnected, long-sequence household tasks often face difficulties with in-context memory, leading to inefficiencies and errors in t…
DaDu-E: Rethinking the Role of Large Language Model in Robotic Computing Pipeline
Wenhao Sun, Sai Hou, Zixuan Wang +6
Performing complex tasks in open environments remains challenging for robots, even when using large language models (LLMs) as the core planner. Many LLM-based planners are ineffici…