24 citations · 25 across the 5 of their papers we have counts for
1 paper · 1 filter
Huan Ma, Yan Zhu, Changqing Zhang +5
Vision-language foundation models have exhibited remarkable success across a multitude of downstream tasks due to their scalability on extensive image-text paired data. However, th…