4 citations · 4 across the 1 of their papers we have counts for
3 papers
cs.CV2023★ 4 cited
Parameter and Computation Efficient Transfer Learning for Vision-Language Pre-trained Models
Qiong Wu, Wei Yu, Yiyi Zhou +3
With ever increasing parameters and computation, vision-language pre-trained (VLP) models exhibit prohibitive expenditure in downstream task adaption. Recent endeavors mainly focus…
cs.CV2023
Approximated Prompt Tuning for Vision-Language Pre-trained Models
Qiong Wu, Shubin Huang, Yiyi Zhou +4
Prompt tuning is a parameter-efficient way to deploy large-scale pre-trained models to downstream tasks by adding task-specific tokens. In terms of vision-language pre-trained (VLP…
cs.CV2023
Adapting Pre-trained Language Models to Vision-Language Tasks via Dynamic Visual Prompting
Shubin Huang, Qiong Wu, Yiyi Zhou +4
Pre-trained language models (PLMs) have played an increasing role in multimedia research. In terms of vision-language (VL) tasks, they often serve as a language encoder and still r…