2 papers
cs.CV2025
Hints of Prompt: Enhancing Visual Representation for Multimodal LLMs in Autonomous Driving
Hao Zhou, Zhanning Gao, Zhili Chen +4
In light of the dynamic nature of autonomous driving environments and stringent safety requirements, general MLLMs combined with CLIP alone often struggle to accurately represent d…
cs.RO2025
NetRoller: Interfacing General and Specialized Models for End-to-End Autonomous Driving
Ren Xin, Hongji Liu, Xiaodong Mei +4
Integrating General Models (GMs) such as Large Language Models (LLMs), with Specialized Models (SMs) in autonomous driving tasks presents a promising approach to mitigating challen…