2 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.CV2021★ 2 cited
Cross-Modal Generative Augmentation for Visual Question Answering
Zixu Wang, Yishu Miao, Lucia Specia
Data augmentation has been shown to effectively improve the performance of multimodal machine learning models. This paper introduces a generative model for data augmentation by lev…
cs.CV2021★ 2 cited
Latent Variable Models for Visual Question Answering
Zixu Wang, Yishu Miao, Lucia Specia
Current work on Visual Question Answering (VQA) explore deterministic approaches conditioned on various types of image and question features. We posit that, in addition to image an…