3 papers
cs.RO2026
LeapBot-WA: World-Anchor Action Models via Predictive Latent Alignments
Pei Liu, Nan Zheng, Lang Zhang +8
World Action Models (WAMs) have emerged as a powerful paradigm for embodied intelligence, yet the prevailing reliance on pixel-level video generation creates a fundamental bottlene…
cs.CV2025
Decoupling Scene Perception and Ego Status: A Multi-Context Fusion Approach for Enhanced Generalization in End-to-End Autonomous Driving
Jiacheng Tang, Mingyue Feng, Jiachao Liu +2
Modular design of planning-oriented autonomous driving has markedly advanced end-to-end systems. However, existing architectures remain constrained by an over-reliance on ego statu…
cs.CV2025
TrajFlow: Multi-modal Motion Prediction via Flow Matching
Qi Yan, Brian Zhang, Yutong Zhang +8
Efficient and accurate motion prediction is crucial for ensuring safety and informed decision-making in autonomous driving, particularly under dynamic real-world conditions that ne…