2 citations · 2 across the 1 of their papers we have counts for
2 papers
cs.CV2025★ 2 cited
A Survey on Vision-Language-Action Models for Autonomous Driving
Sicong Jiang, Zilin Huang, Kangan Qian +17
The rapid progress of multimodal large language models (MLLM) has paved the way for Vision-Language-Action (VLA) paradigms, which integrate visual perception, natural language unde…
cs.LG2025
From Street Views to Urban Science: Discovering Road Safety Factors with Multimodal Large Language Models
Yihong Tang, Ao Qu, Xujing Yu +4
Urban and transportation research has long sought to uncover statistically meaningful relationships between key variables and societal outcomes such as road safety, to generate act…