1 paper
Dingyang Lin, Yingfeng Luo, Chenglong Wang +4
Vision-language models (VLMs) have achieved strong performance on tasks such as visual question answering, yet small image resizes can turn correct answers into errors. We investig…