activity
20152020
most citedAttention-Aware Face Hallucination via Deep Reinforcement Learning

43 citations · 69 across the 6 of their papers we have counts for

collaborators

9 papers

cs.CV20205 cited

Knowledge-Routed Visual Question Reasoning: Challenges for Deep Representation Embedding

Qingxing Cao, Bailin Li, Xiaodan Liang +2

Though beneficial for encouraging the Visual Question Answering (VQA) models to discover the underlying knowledge by exploiting the input-output correlation beyond image and text c…

cs.CL2020

REM-Net: Recursive Erasure Memory Network for Commonsense Evidence Refinement

Yinya Huang, Meng Fang, Xunlin Zhan +3

When answering a question, people often draw upon their rich world knowledge in addition to the particular context. While recent works retrieve supporting facts/evidence from commo…

cs.CV2020

Linguistically Driven Graph Capsule Network for Visual Question Reasoning

Qingxing Cao, Xiaodan Liang, Keze Wang +1

Recently, studies of visual question answering have explored various architectures of end-to-end networks and achieved promising results on both natural and synthetic datasets, whi…

cs.CV20199 cited

Explainable High-order Visual Question Reasoning: A New Benchmark and Knowledge-routed Network

Qingxing Cao, Bailin Li, Xiaodan Liang +1

Explanation and high-order reasoning capabilities are crucial for real-world visual question answering with diverse levels of inference complexity (e.g., what is the dog that is ne…

cs.CV2019

Face Hallucination by Attentive Sequence Optimization with Reinforcement Learning

Yukai Shi, Guanbin Li, Qingxing Cao +2

Face hallucination is a domain-specific super-resolution problem that aims to generate a high-resolution (HR) face image from a low-resolution~(LR) input. In contrast to the existi…

cs.CV2018

Interpretable Visual Question Answering by Reasoning on Dependency Trees

Qingxing Cao, Bailin Li, Xiaodan Liang +1

Collaborative reasoning for understanding image-question pairs is a very critical but underexplored topic in interpretable visual question answering systems. Although very recent s…