2 papers
cs.CV2025
From Prototypes to General Distributions: An Efficient Curriculum for Masked Image Modeling
Jinhong Lin, Cheng-En Wu, Huanran Li +3
Masked Image Modeling (MIM) has emerged as a powerful self-supervised learning paradigm for visual representation learning, enabling models to acquire rich visual representations b…
cs.CV2024
Patch Ranking: Efficient CLIP by Learning to Rank Local Patches
Cheng-En Wu, Jinhong Lin, Yu Hen Hu +1
Contrastive image-text pre-trained models such as CLIP have shown remarkable adaptability to downstream tasks. However, they face challenges due to the high computational requireme…