4 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CV2025★ 3 cited
Multi-Modal Foundation Models for Computational Pathology: A Survey
Dong Li, Guihong Wan, Xintao Wu +7
Foundation models have emerged as a powerful paradigm in computational pathology (CPath), enabling scalable and generalizable analysis of histopathological images. While early deve…
cs.CV2024★ 4 cited
CMFN: Cross-Modal Fusion Network for Irregular Scene Text Recognition
Jinzhi Zheng, Ruyi Ji, Libo Zhang +2
Scene text recognition, as a cross-modal task involving vision and text, is an important research topic in computer vision. Most existing methods use language models to extract sem…