1 paper · 1 filter
Dawei Zhu, Rui Meng, Jiefeng Chen +3
Comprehending long visual documents, where information is distributed across extensive pages of text and visual elements, is a critical but challenging task for modern Vision-Langu…