3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2025
LMAD: Integrated End-to-End Vision-Language Model for Explainable Autonomous Driving
Nan Song, Bozhou Zhang, Xiatian Zhu +2
Large vision-language models (VLMs) have shown promising capabilities in scene understanding, enhancing the explainability of driving behaviors and interactivity with users. Existi…
cs.CV2022★ 3 cited
Few-shot Open-set Recognition Using Background as Unknowns
Nan Song, Chi Zhang, Guosheng Lin
Few-shot open-set recognition aims to classify both seen and novel images given only limited training data of seen classes. The challenge of this task is that the model is required…