3 papers
cs.CV2024
Prompting DirectSAM for Semantic Contour Extraction in Remote Sensing Images
Shiyu Miao, Delong Chen, Fan Liu +4
The Direct Segment Anything Model (DirectSAM) excels in class-agnostic contour extraction. In this paper, we explore its use by applying it to optical remote sensing imagery, where…
cs.CV2024
Making Large Vision Language Models to be Good Few-shot Learners
Fan Liu, Wenwen Cai, Jian Huo +3
Few-shot classification (FSC) is a fundamental yet challenging task in computer vision that involves recognizing novel classes from limited data. While previous methods have focuse…
cs.CV2024
Few-shot Adaptation of Multi-modal Foundation Models: A Survey
Fan Liu, Tianshu Zhang, Wenwen Dai +3
Multi-modal (vision-language) models, such as CLIP, are replacing traditional supervised pre-training models (e.g., ImageNet-based pre-training) as the new generation of visual fou…