78 citations · 80 across the 4 of their papers we have counts for
4 papers · 1 filter
Reasoning with Multi-Structure Commonsense Knowledge in Visual Dialog
Shunyu Zhang, Xiaoze Jiang, Zequn Yang +2
Visual Dialog requires an agent to engage in a conversation with humans grounded in an image. Many studies on Visual Dialog focus on the understanding of the dialog history or the…
A sequential guiding network with attention for image captioning
Daouda Sow, Zengchang Qin, Mouhamed Niasse +1
The recent advances of deep learning in both computer vision (CV) and natural language processing (NLP) provide us a new way of understanding semantics, by which we can deal with m…
Pixel Level Data Augmentation for Semantic Image Segmentation using Generative Adversarial Networks
Shuangting Liu, Jiaqi Zhang, Yuxin Chen +3
Semantic segmentation is one of the basic topics in computer vision, it aims to assign semantic labels to every pixel of an image. Unbalanced semantic label distribution could have…
Data Augmentation in Emotion Classification Using Generative Adversarial Networks
Xinyue Zhu, Yifan Liu, Zengchang Qin +1
It is a difficult task to classify images with multiple class labels using only a small number of labeled examples, especially when the label (class) distribution is imbalanced. Em…