1 paper
Duaa Alim, Mogtaba Alim, Liam Chalcroft
Vision-language models (VLMs) read an image and produce text in a single forward pass, whereas radiologists typically inspect an image several times and consult the literature befo…