Showing cs.ROShow all
3 papers · 1 filter
cs.RO2026
PlayWorld: Learning Robot World Models from Autonomous Play
Tenny Yin, Zhiting Mei, Zhonghe Zheng +8
Action-conditioned video models offer a promising path to building general-purpose robot simulators that can improve directly from data. Yet, despite training on large-scale robot…
cs.RO2025
WoMAP: World Models For Embodied Open-Vocabulary Object Localization
Tenny Yin, Zhiting Mei, Tao Sun +6
Language-instructed active object localization is a critical challenge for robots, requiring efficient exploration of partially observable environments. However, state-of-the-art a…
cs.RO2025
VERDI: VLM-Embedded Reasoning for Autonomous Driving
Bowen Feng, Zhiting Mei, Julian Ost +5
While autonomous driving (AD) stacks struggle with decision making under partial observability and real-world complexity, human drivers are capable of applying commonsense reasonin…