output
20172020
most citedDIDFuse: Deep Image Decomposition for Infrared and Visible Image Fusion

328 citations

13 papers

cs.CV2020★ 27 cited

Counterfactual Samples Synthesizing for Robust Visual Question Answering

Long Chen, Xin Yan, Jun Xiao +3

Despite Visual Question Answering (VQA) has realized impressive progress over the last few years, today's VQA models tend to capture superficial linguistic correlations in the trai…

eess.IV2020★ 328 cited

DIDFuse: Deep Image Decomposition for Infrared and Visible Image Fusion

Zixiang Zhao, Shuang Xu, Chunxia Zhang +3

Infrared and visible image fusion, a hot topic in the field of image processing, aims at obtaining fused images keeping the advantages of source images. This paper proposes a novel…

cs.CV2019★ 2 cited

Adversarial Seeded Sequence Growing for Weakly-Supervised Temporal Action Localization

Chengwei Zhang, Yunlu Xu, Zhanzhan Cheng +4

Temporal action localization is an important yet challenging research topic due to its various applications. Since the frame-level or segment-level annotations of untrimmed videos…

cs.CV2019

REAPS: Towards Better Recognition of Fine-grained Images by Region Attending and Part Sequencing

Peng Zhang, Xinyu Zhu, Zhanzhan Cheng +2

Fine-grained image recognition has been a hot research topic in computer vision due to its various applications. The-state-of-the-art is the part/region-based approaches that first…

cs.CV2019★ 79 cited

Semantic-Guided Multi-Attention Localization for Zero-Shot Learning

Yizhe Zhu, Jianwen Xie, Zhiqiang Tang +2

Zero-shot learning extends the conventional object classification to the unseen class recognition by introducing semantic representations of classes. Existing approaches predominan…

cs.CV2018★ 9 cited

Segregated Temporal Assembly Recurrent Networks for Weakly Supervised Multiple Action Detection

Yunlu Xu, Chengwei Zhang, Zhanzhan Cheng +4

This paper proposes a segregated temporal assembly recurrent (STAR) network for weakly-supervised multiple action detection. The model learns from untrimmed videos with only superv…