2 papers
cs.CV2025
A Survey on Vision-Language-Action Models for Autonomous Driving
Sicong Jiang, Zilin Huang, Kangan Qian +17
The rapid progress of multimodal large language models (MLLM) has paved the way for Vision-Language-Action (VLA) paradigms, which integrate visual perception, natural language unde…
cs.CV2025
POD: Predictive Object Detection with Single-Frame FMCW LiDAR Point Cloud
Yining Shi, Kun Jiang, Xin Zhao +5
LiDAR-based 3D object detection is a fundamental task in the field of autonomous driving. This paper explores the unique advantage of Frequency Modulated Continuous Wave (FMCW) LiD…