Showing cs.CVShow all
2 papers · 1 filter
cs.CV2026
Rethinking Model Selection in VLM Through the Lens of Gromov-Wasserstein Distance
Muyang Li, Yucheng Liu, Jianbo Ma +3
Vision-Language Models (VLMs) have enhanced traditional LLMs with visual capabilities through the integration of vision encoders. While recent works have explored various combinati…
cs.CV2024
Mind the Gap Between Prototypes and Images in Cross-domain Finetuning
Hongduan Tian, Feng Liu, Zhanke Zhou +3
In cross-domain few-shot classification (CFC), recent works mainly focus on adapting a simple transformation head on top of a frozen pre-trained backbone with few labeled data to p…