1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CV2025
Vision-Language Models for Autonomous Driving: CLIP-Based Dynamic Scene Understanding
Mohammed Elhenawy, Huthaifa I. Ashqar, Andry Rakotonirainy +3
Scene understanding is essential for enhancing driver safety, generating human-centric explanations for Automated Vehicle (AV) decisions, and leveraging Artificial Intelligence (AI…
cs.CV2024★ 1 cited
Advancing Object Detection in Transportation with Multimodal Large Language Models (MLLMs): A Comprehensive Review and Empirical Testing
Huthaifa I. Ashqar, Ahmed Jaber, Taqwa I. Alhadidi +1
This study aims to comprehensively review and empirically evaluate the application of multimodal large language models (MLLMs) and Large Vision Models (VLMs) in object detection fo…