2 papers
cs.CV2026
MedicalNarratives: Connecting Medical Vision and Language with Localized Narratives
Wisdom O. Ikezogwo, Kevin Zhang, Mehmet Saygin Seyfioglu +3
Multi-modal models are data hungry. While datasets with natural images are abundant, medical image datasets can not afford the same luxury. To enable representation learning for me…
cs.AI2025
MedBLINK: Probing Basic Perception in Multimodal Language Models for Medicine
Mahtab Bigverdi, Wisdom Ikezogwo, Kevin Zhang +5
Multimodal language models (MLMs) show promise for clinical decision support and diagnostic reasoning, raising the prospect of end-to-end automated medical image interpretation. Ho…