1 paper
Zhongkuan Mao, Wenzhuo Zhao, Xianjie Liu +7
Failures of high-resolution MLLMs are commonly attributed to a visual problem, motivating zooming, cropping, and related visual interventions to recover fine-grained evidence or su…