2 papers
cs.CV2026
Prompt Tuning for CLIP on the Pretrained Manifold
Xi Yang, Yuanrong Xu, Weigang Zhang +3
Prompt tuning introduces learnable prompt vectors that adapt pretrained vision-language models to downstream tasks in a parameter-efficient manner. However, under limited supervisi…
cs.CV2026
Cross-Modal Mapping: Mitigating the Modality Gap for Few-Shot Image Classification
Xi Yang, Pai Peng, Wulin Xie +2
Few-shot image classification remains a critical challenge in the field of computer vision, particularly in data-scarce environments. Existing methods typically rely on pre-trained…