2 papers
cs.CV2024
Towards Efficient Vision-Language Tuning: More Information Density, More Generalizability
Tianxiang Hao, Mengyao Lyu, Hui Chen +4
With the advancement of large pre-trained vision-language models, effectively transferring the knowledge embedded within these foundational models to downstream tasks has become a…
cs.CV2024
Quantized Prompt for Efficient Generalization of Vision-Language Models
Tianxiang Hao, Xiaohan Ding, Juexiao Feng +3
In the past few years, large-scale pre-trained vision-language models like CLIP have achieved tremendous success in various fields. Naturally, how to transfer the rich knowledge in…