33 citations · 52 across the 4 of their papers we have counts for
1 paper · 1 filter
Hexiang Hu, Yi Luan, Yang Chen +5
Large-scale multi-modal pre-training models such as CLIP and PaLI exhibit strong generalization on various visual domains and tasks. However, existing image classification benchmar…