Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Qwen-RobotWorld Technical Report: Unifying Embodied World Modeling through Language-Conditioned Video Generation
Jie Zhang, Xiaoyue Chen, Anzhe Chen +36
We introduce Qwen-RobotWorld, a language-conditioned video world model for embodied intelligence. With natural language as a unified action interface, it predicts physically ground…
cs.CV2026
GraphWorld: Long-Horizon Planning with World Models for End-to-End Autonomous Driving
Ziying Song, Caiyan Jia, Lin Liu +8
End-to-end autonomous driving has made significant progress by unifying perception, prediction, and planning within a single learning framework, achieving strong performance in sho…