31 citations · 32 across the 3 of their papers we have counts for
3 papers
cs.LG2024
Understanding Retrieval-Augmented Task Adaptation for Vision-Language Models
Yifei Ming, Yixuan Li
Pre-trained contrastive vision-language models have demonstrated remarkable performance across a wide range of tasks. However, they often struggle on fine-trained datasets with cat…
cs.CV2022★ 31 cited
Delving into Out-of-Distribution Detection with Vision-Language Representations
Yifei Ming, Ziyang Cai, Jiuxiang Gu +3
Recognizing out-of-distribution (OOD) samples is critical for machine learning systems deployed in the open world. The vast majority of OOD detection methods are driven by a single…
cs.CV2022★ 1 cited
Are Vision Transformers Robust to Spurious Correlations?
Soumya Suvra Ghosal, Yifei Ming, Yixuan Li
Deep neural networks may be susceptible to learning spurious correlations that hold on average but not in atypical test samples. As with the recent emergence of vision transformer…