3 papers
cs.CV2026
EyeMVP: OCT-Informed Fundus Representation Learning via Paired CFP--OCT Pretraining
Zhuo Deng, Ruiheng Zhang, Ziheng Zhang +23
Color fundus photography (CFP) is the mainstay of large-scale retinal screening, but its diagnostic capacity is limited by the lack of depth-resolved structure, which optical coher…
cs.CV2024
Universal Medical Image Representation Learning with Compositional Decoders
Kaini Wang, Ling Yang, Siping Zhou +4
Visual-language models have advanced the development of universal models, yet their application in medical imaging remains constrained by specific functional requirements and the l…
cs.CV2024
TSdetector: Temporal-Spatial Self-correction Collaborative Learning for Colonoscopy Video Detection
Kaini Wang, Haolin Wang, Guang-Quan Zhou +4
CNN-based object detection models that strike a balance between performance and speed have been gradually used in polyp detection tasks. Nevertheless, accurately locating polyps wi…