chain-of-thought reasoning 1embodied ai 1foundation models 1pixel grounding 1visual language navigation 1
From the 1 of 8 linked papers with an AI index.
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
ABot-N1: Toward a General Visual Language Navigation Foundation Model
Ruiyan Gong, Yingnan Guo, Junjun Hu +44
The paper presents ABot-N1, a visual‑language navigation foundation model that separates high‑level reasoning from low‑level control via a slow‑fast architecture and pixel‑based go…
cs.CV2026
Grounded Forcing: Bridging Time-Independent Semantics and Proximal Dynamics in Autoregressive Video Synthesis
Jintao Chen, Chengyu Bai, Junjun Hu +2
Autoregressive video synthesis offers a promising pathway for infinite-horizon generation but is fundamentally hindered by three intertwined challenges: semantic forgetting from co…
cs.CV2026
AstraNav-World: World Model for Foresight Control and Consistency
Jintao Chen, Junjun Hu, Haochen Bai +11
Embodied navigation in open, dynamic environments demands accurate foresight of how the world will evolve and how actions will unfold over time. We propose AstraNav-World, an end-t…