4 papers
AnyMatch: Supercharging Universal Multi-Modal Image Matching with Large-Scale Single-View Images
Meng Yang, Zizhuo Li, Linfeng Tang +2
Multi-modal image matching is essential for visual localization and multi-sensor fusion, but it is hindered by the scarcity of large-scale training data with precise geometric anno…
DistillMatch: Leveraging Knowledge Distillation from Vision Foundation Model for Multimodal Image Matching
Meng Yang, Fan Fan, Zizhuo Li +3
Multimodal image matching seeks pixel-level correspondences between images of different modalities, crucial for cross-modal perception, fusion and analysis. However, the significan…
A Fully Open and Generalizable Foundation Model for Ultrasound Clinical Applications
Hongyuan Zhang, Yuheng Wu, Mingyang Zhao +22
Artificial intelligence (AI) that can effectively learn ultrasound representations by integrating multi-source data holds significant promise for advancing clinical care. However,…
Benchmarking Chest X-ray Diagnosis Models Across Multinational Datasets
Qinmei Xu, Yiheng Li, Xianghao Zhan +10
Foundation models leveraging vision-language pretraining have shown promise in chest X-ray (CXR) interpretation, yet their real-world performance across diverse populations and dia…