From the 1 of 1 linked paper with an AI index.
1 paper
Jiaang Li, Chengzu Li, Zhaochong An +4
The paper investigates why multimodal large language models often ignore visual evidence, using image reconstruction and a new benchmark (WhatIfVis) to measure how well models bala…