3 papers
cs.DC2025
CacheFL: Privacy-Preserving and Efficient Federated Cache Model Fine-Tuning for Vision-Language Models
Mengjun Yi, Hanwen Zhang, Hui Dou +2
Large pre-trained Vision-Language Models (VLMs), such as Contrastive Language-Image Pre-training (CLIP), have exhibited remarkable zero-shot performance across various image classi…
cs.CV2025
Interactive Instance Annotation with Siamese Networks
Xiang Xu, Ruotong Li, Mengjun Yi +3
Annotating instance masks is time-consuming and labor-intensive. A promising solution is to predict contours using a deep learning model and then allow users to refine them. Howeve…
cs.LG2024
Explaining Model Overfitting in CNNs via GMM Clustering
Hui Dou, Xinyu Mu, Mengjun Yi +3
Convolutional Neural Networks (CNNs) have demonstrated remarkable prowess in the field of computer vision. However, their opaque decision-making processes pose significant challeng…