Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
CoCa-CXR: Contrastive Captioners Learn Strong Temporal Structures for Chest X-Ray Vision-Language Understanding
Yixiong Chen, Shawn Xu, Andrew Sellergren +6
Vision-language models have proven to be of great benefit for medical image analysis since they learn rich semantics from both images and reports. Prior efforts have focused on bet…
cs.CV2025
PolyPath: Adapting a Large Multimodal Model for Multi-slide Pathology Report Generation
Faruk Ahmed, Lin Yang, Tiam Jaroensri +11
The interpretation of histopathology cases underlies many important diagnostic and treatment decisions in medicine. Notably, this process typically requires pathologists to integra…