1 paper
Haoyu Zuo, Yibo Yan, Xin Zou +4
Multi-vector vision-language retrievers enable fine-grained Visual Document Retrieval (VDR) through late interaction, but storing and scoring hundreds of visual patch embeddings pe…