72 citations · 99 across the 4 of their papers we have counts for
4 papers
VisionFM: a Multi-Modal Multi-Task Vision Foundation Model for Generalist Ophthalmic Artificial Intelligence
Jianing Qiu, Jian Wu, Hao Wei +39
We present VisionFM, a foundation model pre-trained with 3.4 million ophthalmic images from 560,457 individuals, covering a broad range of ophthalmic diseases, modalities, imaging…
Linguistic Query-Guided Mask Generation for Referring Image Segmentation
Zhichao Wei, Xiaohao Chen, Mingqiang Chen +1
Referring image segmentation aims to segment the image region of interest according to the given language expression, which is a typical multi-modal task. Existing methods either a…
GB-CosFace: Rethinking Softmax-based Face Recognition from the Perspective of Open Set Classification
Lizhe Liu, Mingqiang Chen, Xiaohao Chen +2
State-of-the-art face recognition methods typically take the multi-classification pipeline and adopt the softmax-based loss for optimization. Although these methods have achieved g…
Batch DropBlock Network for Person Re-identification and Beyond
Zuozhuo Dai, Mingqiang Chen, Xiaodong Gu +2
Since the person re-identification task often suffers from the problem of pose changes and occlusions, some attentive local features are often suppressed when training CNNs. In thi…