Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Fine-Grained VLM Fine-tuning via Latent Hierarchical Adapter Learning
Yumiao Zhao, Bo Jiang, Yuhe Ding +3
Adapter-based approaches have garnered attention for fine-tuning pre-trained Vision-Language Models (VLMs) on few-shot classification tasks. These methods strive to develop a light…
cs.CV2025
Harmonizing and Merging Source Models for CLIP-based Domain Generalization
Yuhe Ding, Jian Liang, Bo Jiang +3
CLIP-based domain generalization aims to improve model generalization to unseen domains by leveraging the powerful zero-shot classification capabilities of CLIP and multiple source…
cs.CV2024
Learning to Rank Pre-trained Vision-Language Models for Downstream Tasks
Yuhe Ding, Bo Jiang, Aihua Zheng +2
Vision language models (VLMs) like CLIP show stellar zero-shot capability on classification benchmarks. However, selecting the VLM with the highest performance on the unlabeled dow…