activity
20242026
collaborators

5 papers

cs.AI2026

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

Siddharth Damodharan, Radhika Gupta, Ali Alshami +2

Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as scene understanding, decision…

cs.LG2026

Adapting Feature Attenuation to NLP

Tianshuo Yang, Ryan Rabinowitz, Terrance E. Boult +1

Transformer classifiers such as BERT deliver impressive closed-set accuracy, yet they remain brittle when confronted with inputs from unseen categories--a common scenario for deplo…

cs.CV2025

2COOOL: 2nd Workshop on the Challenge Of Out-Of-Label Hazards in Autonomous Driving

Ali K. AlShami, Ryan Rabinowitz, Maged Shoman +9

As the computer vision community advances autonomous driving algorithms, integrating vision-based insights with sensor data remains essential for improving perception, decision mak…

cs.CV2025

SMART-Vision: Survey of Modern Action Recognition Techniques in Vision

Ali K. AlShami, Ryan Rabinowitz, Khang Lam +4

Human Action Recognition (HAR) is a challenging domain in computer vision, involving recognizing complex patterns by analyzing the spatiotemporal dynamics of individuals' movements…

cs.CV2024

COOOL: Challenge Of Out-Of-Label A Novel Benchmark for Autonomous Driving

Ali K. AlShami, Ananya Kalita, Ryan Rabinowitz +4

As the Computer Vision community rapidly develops and advances algorithms for autonomous driving systems, the goal of safer and more efficient autonomous transportation is becoming…