2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.CV2023★ 1 cited
From Association to Generation: Text-only Captioning by Unsupervised Cross-modal Mapping
Junyang Wang, Ming Yan, Yi Zhang +1
With the development of Vision-Language Pre-training Models (VLPMs) represented by CLIP and ALIGN, significant breakthroughs have been achieved for association-based visual tasks s…
cs.CV2022★ 2 cited
Counterfactually Measuring and Eliminating Social Bias in Vision-Language Pre-training Models
Yi Zhang, Junyang Wang, Jitao Sang
Vision-Language Pre-training (VLP) models have achieved state-of-the-art performance in numerous cross-modal tasks. Since they are optimized to capture the statistical properties o…