3 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Guiming Cao, Kaize Shi, Hong Fu +2
Pre-trained Vision-Language (V-L) models set the benchmark for generalization to downstream tasks among the noteworthy contenders. Many characteristics of the V-L model have been e…