2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Qiying Yu, Quan Sun, Xiaosong Zhang +5
Large multimodal models demonstrate remarkable generalist ability to perform diverse multimodal tasks in a zero-shot manner. Large-scale web-based image-text pairs contribute funda…