2 papers
cs.RO2026
NaviDriveVLM: Decoupling High-Level Reasoning and Motion Planning for Autonomous Driving
Ximeng Tao, Pardis Taghavi, Dimitar Filev +2
Vision-language models (VLMs) have emerged as a promising direction for end-to-end autonomous driving (AD) by jointly modeling visual observations, driving context, and language-ba…
cs.CV2026
Toward Unified Multimodal Representation Learning for Autonomous Driving
Ximeng Tao, Dimitar Filev, Gaurav Pandey
Contrastive Language-Image Pre-training (CLIP) has shown impressive performance in aligning visual and textual representations. Recent studies have extended this paradigm to 3D vis…