3 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2020★ 3 cited
Overcoming Language Priors with Self-supervised Learning for Visual Question Answering
Xi Zhu, Zhendong Mao, Chunxiao Liu +3
Most Visual Question Answering (VQA) models suffer from the language prior problem, which is caused by inherent data biases. Specifically, VQA models tend to answer questions (e.g.…
cs.SD2020★ 3 cited
Audio-visual Speech Separation with Adversarially Disentangled Visual Representation
Peng Zhang, Jiaming Xu, Jing shi +2
Speech separation aims to separate individual voice from an audio mixture of multiple simultaneous talkers. Although audio-only approaches achieve satisfactory performance, they bu…