Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
Exploring scalable medical image encoders beyond text supervision
Fernando Pérez-GarcÃa, Harshita Sharma, Sam Bond-Taylor +12
Language-supervised pre-training has proven to be a valuable method for extracting semantically meaningful features from images, serving as a foundational element in multimodal sys…
cs.CV2024
MAIRA-Seg: Enhancing Radiology Report Generation with Segmentation-Aware Multimodal Large Language Models
Harshita Sharma, Valentina Salvatelli, Shaury Srivastav +13
There is growing interest in applying AI to radiology report generation, particularly for chest X-rays (CXRs). This paper investigates whether incorporating pixel-level information…