2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Xupeng Chen, Binbin Shi, Chenqian Le +5
Vision-language models (VLMs) are increasingly applied to medical visual question answering (Med-VQA), yet whether they can \emph{localize} the evidence behind their answers---a pr…