1 paper
Chengsong You, Qiyi Jiang, Junwei Zhou +12
Visual document retrieval has advanced by encoding page screenshots with vision-language models, bypassing OCR pipelines. However, existing methods remain page-centric, misaligned…