8 papers
StepAL: Step-aware Active Learning for Cataract Surgical Videos
Nisarg A. Shah, Bardia Safaei, Shameema Sikder +2
Active learning (AL) can reduce annotation costs in surgical video analysis while maintaining model performance. However, traditional AL methods, developed for images or short vide…
Certainty and Uncertainty Guided Active Domain Adaptation
Bardia Safaei, Vibashan VS, Vishal M. Patel
Active Domain Adaptation (ADA) adapts models to target domains by selectively labeling a few target samples. Existing ADA methods prioritize uncertain samples but overlook confiden…
On the Performance of Unmanned Aerial Vehicles with MIMO VLC
Hosein Zarini, Amir Mohammadisarab, Maryam Farajzadeh Dehkordi +5
This paper centers around a multiple-input-multiple-output (MIMO) visible light communication (VLC) system, where an unmanned aerial vehicle (UAV) benefits from a light emitting di…
Filter Images First, Generate Instructions Later: Pre-Instruction Data Selection for Visual Instruction Tuning
Bardia Safaei, Faizan Siddiqui, Jiacong Xu +2
Visual instruction tuning (VIT) for large vision-language models (LVLMs) requires training on expansive datasets of image-instruction pairs, which can be costly. Recent efforts in…
Towards Zero-Shot Anomaly Detection and Reasoning with Multimodal Large Language Models
Jiacong Xu, Shao-Yuan Lo, Bardia Safaei +2
Zero-Shot Anomaly Detection (ZSAD) is an emerging AD paradigm. Unlike the traditional unsupervised AD setting that requires a large number of normal samples to train a model, ZSAD…
Active Learning for Vision-Language Models
Bardia Safaei, Vishal M. Patel
Pre-trained vision-language models (VLMs) like CLIP have demonstrated impressive zero-shot performance on a wide range of downstream computer vision tasks. However, there still exi…