2 papers
cs.CV2025
MMCLIP: Cross-modal Attention Masked Modelling for Medical Language-Image Pre-Training
Biao Wu, Yutong Xie, Zeyu Zhang +4
Vision-and-language pretraining (VLP) in the medical field utilizes contrastive learning on image-text pairs to achieve effective transfer across tasks. Yet, current VLP approaches…
cs.CV2024
Act Like a Radiologist: Radiology Report Generation across Anatomical Regions
Qi Chen, Yutong Xie, Biao Wu +5
Automating radiology report generation can ease the reporting workload for radiologists. However, existing works focus mainly on the chest area due to the limited availability of p…