Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
On the Robustness of Medical Vision-Language Models: Are they Truly Generalizable?
Raza Imam, Rufael Marew, Mohammad Yaqub
Medical Vision-Language Models (MVLMs) have achieved par excellence generalization in medical image analysis, yet their performance under noisy, corrupted conditions remains largel…
cs.CV2025
CLIP meets DINO for Tuning Zero-Shot Classifier using Unlabeled Image Collections
Mohamed Fazli Imam, Rufael Fedaku Marew, Jameel Hassan +3
In the era of foundation models, CLIP has emerged as a powerful tool for aligning text & visual modalities into a common embedding space. However, the alignment objective used to t…