4 papers
MMBU: A Massive Multi-modal Biomedical Understanding Benchmark to Probe the Perception Capabilities of Vision-Language Models
Ryan D'Cunha, Alejandro Lozano, Xiaoxiao Sun +17
Vision and language models (VLMs) hold immense promise to transform biomedical imaging workflows, from detecting lesions in chest X-rays to profiling cellular features in microscop…
The 8th AI City Challenge
Shuo Wang, David C. Anastasiu, Zheng Tang +21
The eighth AI City Challenge highlighted the convergence of computer vision and artificial intelligence in areas like retail, warehouse settings, and Intelligent Traffic Systems (I…
Open-Set Facial Expression Recognition
Yuhang Zhang, Yue Yao, Xuannan Liu +3
Facial expression recognition (FER) models are typically trained on datasets with a fixed number of seven basic classes. However, recent research works point out that there are far…
Alice Benchmarks: Connecting Real World Re-Identification with the Synthetic
Xiaoxiao Sun, Yue Yao, Shengjin Wang +2
For object re-identification (re-ID), learning from synthetic data has become a promising strategy to cheaply acquire large-scale annotated datasets and effective models, with few…