2 papers
cs.CV2025
TB-Bench: Training and Testing Multi-Modal AI for Understanding Spatio-Temporal Traffic Behaviors from Dashcam Images/Videos
Korawat Charoenpitaks, Van-Quang Nguyen, Masanori Suganuma +4
The application of Multi-modal Large Language Models (MLLMs) in Autonomous Driving (AD) faces significant challenges due to their limited training on traffic-specific data and the…
cs.CV2024
Exploring the Potential of Multi-Modal AI for Driving Hazard Prediction
Korawat Charoenpitaks, Van-Quang Nguyen, Masanori Suganuma +3
This paper addresses the problem of predicting hazards that drivers may encounter while driving a car. We formulate it as a task of anticipating impending accidents using a single…