1 paper
Wenbo Hu, Jia-Chen Gu, Zi-Yi Dou +4
Existing multimodal retrieval benchmarks primarily focus on evaluating whether models can retrieve and utilize external textual knowledge for question answering. However, there are…