1 citations · 1 across the 1 of their papers we have counts for
3 papers
cs.CV2023★ 1 cited
UniBoost: Unsupervised Unimodal Pre-training for Boosting Zero-shot Vision-Language Tasks
Yanan Sun, Zihan Zhong, Qi Fan +2
Large-scale joint training of multimodal models, e.g., CLIP, have demonstrated great performance in many vision-language tasks. However, image-text pairs for pre-training are restr…
cs.CL2023
Tailoring Instructions to Student's Learning Levels Boosts Knowledge Distillation
Yuxin Ren, Zihan Zhong, Xingjian Shi +3
It has been commonly observed that a teacher model with superior performance does not necessarily result in a stronger student, highlighting a discrepancy between current teacher t…
cs.CV2023
Towards Arbitrary Text-driven Image Manipulation via Space Alignment
Yunpeng Bai, Zihan Zhong, Chao Dong +3
The recent GAN inversion methods have been able to successfully invert the real image input to the corresponding editable latent code in StyleGAN. By combining with the language-vi…