5 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2024
DARA: Domain- and Relation-aware Adapters Make Parameter-efficient Tuning for Visual Grounding
Ting Liu, Xuyang Liu, Siteng Huang +5
Visual grounding (VG) is a challenging task to localize an object in an image based on a textual description. Recent surge in the scale of VG models has substantially improved perf…
cs.CV2024★ 2 cited
Prompt-based Distribution Alignment for Unsupervised Domain Adaptation
Shuanghao Bai, Min Zhang, Wanqi Zhou +4
Recently, despite the unprecedented success of large pre-trained visual-language models (VLMs) on a wide range of downstream tasks, the real-world unsupervised domain adaptation (U…
cs.CV2022★ 5 cited
Tree Structure-Aware Few-Shot Image Classification via Hierarchical Aggregation
Min Zhang, Siteng Huang, Wenbin Li +1
In this paper, we mainly focus on the problem of how to learn additional feature representations for few-shot image classification through pretext tasks (e.g., rotation or color pe…