14 citations · 15 across the 5 of their papers we have counts for
Showing cs.CVShow all
3 papers · 1 filter
cs.CV2024
Improving Data Augmentation for Robust Visual Question Answering with Effective Curriculum Learning
Yuhang Zheng, Zhen Wang, Long Chen
Being widely used in learning unbiased visual question answering (VQA) models, Data Augmentation (DA) helps mitigate language biases by generating extra training samples beyond the…
cs.CV2023
Delving into Shape-aware Zero-shot Semantic Segmentation
Xinyu Liu, Beiwen Tian, Zhen Wang +5
Thanks to the impressive progress of large-scale vision-language pretraining, recent recognition models can classify arbitrary objects in a zero-shot and open-set manner, with a su…
cs.CV2022
Explicit Image Caption Editing
Zhen Wang, Long Chen, Wenbo Ma +4
Given an image and a reference caption, the image caption editing task aims to correct the misalignment errors and generate a refined caption. However, all existing caption editing…