1 paper
Jiahao Huo, Yu Huang, Yibo Yan +7
Although Multimodal Large Language Models (MLLMs) have shown remarkable potential in Visual Document Retrieval (VDR) through generating high-quality multi-vector embeddings, the su…