3 citations · 6 across the 2 of their papers we have counts for
2 papers
cs.CV2020★ 3 cited
Learning content and context with language bias for Visual Question Answering
Chao Yang, Su Feng, Dongsheng Li +3
Visual Question Answering (VQA) is a challenging multimodal task to answer questions about an image. Many works concentrate on how to reduce language bias which makes models answer…
cs.CV2020★ 3 cited
PPGN: Phrase-Guided Proposal Generation Network For Referring Expression Comprehension
Chao Yang, Guoqing Wang, Dongsheng Li +3
Reference expression comprehension (REC) aims to find the location that the phrase refer to in a given image. Proposal generation and proposal representation are two effective tech…