Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
ABot-World-0: Infinite Interactive World Rollout on a Single Desktop GPU
Fan Jiang, Zhaoxu Sun, Mengchao Wang +38
We present ABot-World-0, an action-conditioned video world model for real-time, long-horizon closed-loop interaction, supported by a multi-source data infrastructure spanning AAA g…
cs.CV2026
ReflectVLN: Training Vision-Language Navigation Agents with Reflective Reasoning
Jiahang Wang, Yirong Yang, Yanqing Zhu +4
Existing vision-language navigation methods often couple a VLM with waypoint decoders to produce multi-step action plans, but they typically lack an explicit closed-loop mechanism…