1 paper · 1 filter
Yongxin Shi, Jiapeng Wang, Zeyu Shan +3
Recent multimodal large language models (MLLMs) still struggle with long document understanding due to two fundamental challenges: information interference from abundant irrelevant…