6 citations · 8 across the 7 of their papers we have counts for
6 papers · 1 filter
Chinese Character Recognition with Radical-Structured Stroke Trees
Haiyang Yu, Jingye Chen, Bin Li +1
The flourishing blossom of deep learning has witnessed the rapid development of Chinese character recognition. However, it remains a great challenge that the characters for testing…
Compositional Scene Modeling with Global Object-Centric Representations
Tonglin Chen, Bin Li, Zhimeng Shen +1
The appearance of the same object may vary in different scene images due to perspectives and occlusions between objects. Humans can easily identify the same object, even if occlusi…
Zero-Shot Chinese Character Recognition with Stroke-Level Decomposition
Jingye Chen, Bin Li, Xiangyang Xue
Chinese character recognition has attracted much research interest due to its wide applications. Although it has been studied for many years, some issues in this field have not bee…
A Generic Object Re-identification System for Short Videos
Tairu Qiu, Guanxian Chen, Zhongang Qi +3
Short video applications like TikTok and Kwai have been a great hit recently. In order to meet the increasing demands and take full advantage of visual information in short videos,…
VL-BERT: Pre-training of Generic Visual-Linguistic Representations
Weijie Su, Xizhou Zhu, Yue Cao +4
We introduce a new pre-trainable generic representation for visual-linguistic tasks, called Visual-Linguistic BERT (VL-BERT for short). VL-BERT adopts the simple yet powerful Trans…
Question Guided Modular Routing Networks for Visual Question Answering
Yanze Wu, Qiang Sun, Jianqi Ma +4
This paper studies the task of Visual Question Answering (VQA), which is topical in Multimedia community recently. Particularly, we explore two critical research problems existed i…