75 citations · 165 across the 10 of their papers we have counts for
13 papers · 1 filter
Language-guided Navigation via Cross-Modal Grounding and Alternate Adversarial Learning
Weixia Zhang, Chao Ma, Qi Wu +1
The emerging vision-and-language navigation (VLN) problem aims at learning to navigate an agent to the target location in unseen photo-realistic environments according to the given…
Interpretable Neural Computation for Real-World Compositional Visual Question Answering
Ruixue Tang, Chao Ma
There are two main lines of research on visual question answering (VQA): compositional model with explicit multi-hop reasoning, and monolithic network with implicit reasoning in th…
Rethinking Image Deraining via Rain Streaks and Vapors
Yinglong Wang, Yibing Song, Chao Ma +1
Single image deraining regards an input image as a fusion of a background image, a transmission map, rain streaks, and atmosphere light. While advanced models are proposed for imag…
Robust Tracking against Adversarial Attacks
Shuai Jia, Chao Ma, Yibing Song +1
While deep convolutional neural networks (CNNs) are vulnerable to adversarial attacks, considerably few efforts have been paid to construct robust deep tracking algorithms against…
Unsupervised Deep Representation Learning for Real-Time Tracking
Ning Wang, Wengang Zhou, Yibing Song +3
The advancement of visual tracking has continuously been brought by deep learning models. Typically, supervised learning is employed to train these models with expensive labeled da…
Semantic Equivalent Adversarial Data Augmentation for Visual Question Answering
Ruixue Tang, Chao Ma, Wei Emma Zhang +2
Visual Question Answering (VQA) has achieved great success thanks to the fast development of deep neural networks (DNN). On the other hand, the data augmentation, as one of the maj…