10 citations · 18 across the 6 of their papers we have counts for
1 paper · 1 filter
Yunxin Li, Baotian Hu, Yuxin Ding +2
Pretrained Vision-Language Models (VLMs) have achieved remarkable performance in image retrieval from text. However, their performance drops drastically when confronted with lingui…