Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
A Shared Encoder Approach to Multimodal Representation Learning
Shuvendu Roy, Franklin Ogidi, Ali Etemad +2
Multimodal representation learning has demonstrated remarkable potential in enabling models to process and integrate diverse data modalities, such as text and images, for improved…
cs.CV2024
Benchmarking Vision-Language Contrastive Methods for Medical Representation Learning
Shuvendu Roy, Yasaman Parhizkar, Franklin Ogidi +5
We perform a comprehensive benchmarking of contrastive frameworks for learning multimodal representations in the medical domain. Through this study, we aim to answer the following…