5 papers
AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding
Siddharth Damodharan, Radhika Gupta, Ali Alshami +2
Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene understanding, decision…
Adapting Feature Attenuation to NLP
Tianshuo Yang, Ryan Rabinowitz, Terrance E. Boult +1
Transformer classifiers such as BERT deliver impressive closed-set accuracy, yet they remain brittle when confronted with inputs from unseen categories--a common scenario for deplo…
2COOOL: 2nd Workshop on the Challenge Of Out-Of-Label Hazards in Autonomous Driving
Ali K. AlShami, Ryan Rabinowitz, Maged Shoman +9
As the computer vision community advances autonomous driving algorithms, integrating vision-based insights with sensor data remains essential for improving perception, decision mak…
SMART-Vision: Survey of Modern Action Recognition Techniques in Vision
Ali K. AlShami, Ryan Rabinowitz, Khang Lam +4
Human Action Recognition (HAR) is a challenging domain in computer vision, involving recognizing complex patterns by analyzing the spatiotemporal dynamics of individuals' movements…
COOOL: Challenge Of Out-Of-Label A Novel Benchmark for Autonomous Driving
Ali K. AlShami, Ananya Kalita, Ryan Rabinowitz +4
As the Computer Vision community rapidly develops and advances algorithms for autonomous driving systems, the goal of safer and more efficient autonomous transportation is becoming…