evidence routing 1high-resolution visual question answering 1memory efficiency 1multimodal large language models 1single-pass inference 1
From the 1 of 12 linked papers with an AI index.
Showing cs.ROShow all
3 papers · 1 filter
cs.RO2026
Pelican-VLA 0.5: Attending Before Acting Benefits Generalization
Zeyuan Ding, Wenhai Liu, Yang Xu +6
In this report, we present Pelican-VLA 0.5, a unified VLA model that integrates vision-language understanding, future-frame generation, and action prediction within a single archit…
cs.RO2026
Robo-Cortex: A Self-Evolving Embodied Agent via Dual-Grain Cognitive Memory and Autonomous Knowledge Induction
Nga Teng Chan, Yi Zhang, Yechi Liu +9
The ability to navigate and interact with complex environments is central to real-world embodied agents, yet navigation in unseen environments remains challenging due to "experient…
cs.RO2025
WoW: Towards a World omniscient World model Through Embodied Interaction
Xiaowei Chi, Peidong Jia, Chun-Kai Fan +33
Humans develop an understanding of intuitive physics through active interaction with the world. This approach is in stark contrast to current video models, such as Sora, which rely…