228 citations · 314 across the 11 of their papers we have counts for
18 papers · 1 filter
Learning Conditional Attributes for Compositional Zero-Shot Learning
Qingsheng Wang, Lingqiao Liu, Chenchen Jing +4
Compositional Zero-Shot Learning (CZSL) aims to train models to recognize novel compositional concepts based on learned concepts such as attribute-object combinations. One of the c…
Proposal-free One-stage Referring Expression via Grid-Word Cross-Attention
Wei Suo, Mengyang Sun, Peng Wang +1
Referring Expression Comprehension (REC) has become one of the most important tasks in visual reasoning, since it is an essential step for many vision-and-language tasks such as vi…
Chop Chop BERT: Visual Question Answering by Chopping VisualBERT's Heads
Chenyu Gao, Qi Zhu, Peng Wang +1
Vision-and-Language (VL) pre-training has shown great potential on many related downstream tasks, such as Visual Question Answering (VQA), one of the most popular problems in the V…
An Adversarial Human Pose Estimation Network Injected with Graph Structure
Lei Tian, Guoqiang Liang, Peng Wang +1
Because of the invisible human keypoints in images caused by illumination, occlusion and overlap, it is likely to produce unreasonable human pose prediction for most of the current…
Simple is not Easy: A Simple Strong Baseline for TextVQA and TextCaps
Qi Zhu, Chenyu Gao, Peng Wang +1
Texts appearing in daily scenes that can be recognized by OCR (Optical Character Recognition) tools contain significant information, such as street name, product brand and prices.…
Hyperspectral Classification Based on Lightweight 3-D-CNN With Transfer Learning
Haokui Zhang, Ying Li, Yenan Jiang +3
Recently, hyperspectral image (HSI) classification approaches based on deep learning (DL) models have been proposed and shown promising performance. However, because of very limite…