Showing cs.CVShow all
2 papers · 1 filter
cs.CV2025
One Last Attention for Your Vision-Language Model
Liang Chen, Ghazi Shazan Ahmad, Tianjun Yao +2
Pretrained vision-language models (VLMs), such as CLIP, achieve remarkable zero-shot performance, yet their downstream potential hinges on effective fine-tuning. Most adaptation me…
cs.CV2024
A Causal Inspired Early-Branching Structure for Domain Generalization
Liang Chen, Yong Zhang, Yibing Song +2
Learning domain-invariant semantic representations is crucial for achieving domain generalization (DG), where a model is required to perform well on unseen target domains. One crit…