embodied ai 2benchmark 1chain-of-thought reasoning 1foundation models 1long-horizon tasks 1multi-modal memory 1pixel grounding 1robotic agent operating system 1visual language navigation 1
From the 2 of 3 linked papers with an AI index.
3 papers
cs.AI2026
ABot-AgentOS: A General Robotic Agent OS with Lifelong Multi-modal Memory
Jiayi Tian, Shiao Liu, Yuting Xu +31
The paper introduces ABot-AgentOS, a general operating system layer for robotic agents that adds deliberative planning, multi‑modal memory, verification, and cloud‑edge collaborati…
cs.CV2026
ABot-N1: Toward a General Visual Language Navigation Foundation Model
Ruiyan Gong, Yingnan Guo, Junjun Hu +44
The paper presents ABot-N1, a visual‑language navigation foundation model that separates high‑level reasoning from low‑level control via a slow‑fast architecture and pixel‑based go…
cs.RO2026
ABot-N0: Technical Report on the VLA Foundation Model for Versatile Embodied Navigation
Zedong Chu, Shichao Xie, Xiaolong Wu +41
Embodied navigation has long been fragmented by task-specific architectures. We introduce ABot-N0, a unified Vision-Language-Action (VLA) foundation model that achieves a ``Grand U…