3 citations · 4 across the 12 of their papers we have counts for
1 paper · 1 filter
Hairui Ren, Fan Tang, He Zhao +3
Fine-tuning vision-language models (VLMs) with large amounts of unlabeled data has recently garnered significant interest. However, a key challenge remains the lack of high-quality…