47 citations · 160 across the 18 of their papers we have counts for
22 papers
Multi-Task and Multi-Modal Learning for RGB Dynamic Gesture Recognition
Dinghao Fan, Hengjie Lu, Shugong Xu +1
Gesture recognition is getting more and more popular due to various application possibilities in human-machine interaction. Existing multi-modal gesture recognition systems take mu…
IFR: Iterative Fusion Based Recognizer For Low Quality Scene Text Recognition
Zhiwei Jia, Shugong Xu, Shiyi Mu +3
Although recent works based on deep learning have made progress in improving recognition accuracy on scene text recognition, how to handle low-quality text images in end-to-end dee…
SGTBN: Generating Dense Depth Maps from Single-Line LiDAR
Hengjie Lu, Shugong Xu, Shan Cao
Depth completion aims to generate a dense depth map from the sparse depth map and aligned RGB image. However, current depth completion methods use extremely expensive 64-line LiDAR…
Arbitrary-Shaped Text Detection withAdaptive Text Region Representation
Xiufeng Jiang, Shugong Xu, Shunqing Zhang +1
Text detection/localization, as an important task in computer vision, has witnessed substantialadvancements in methodology and performance with convolutional neural networks. Howev…
Tracking Based Semi-Automatic Annotation for Scene Text Videos
Jiajun Zhu, Xiufeng Jiang, Zhiwei Jia +2
Recently, video scene text detection has received increasing attention due to its comprehensive applications. However, the lack of annotated scene text video datasets has become on…
A Dataset and Benchmark Towards Multi-Modal Face Anti-Spoofing Under Surveillance Scenarios
Xudong Chen, Shugong Xu, Qiaobin Ji +1
Face Anti-spoofing (FAS) is a challenging problem due to complex serving scenarios and diverse face presentation attack patterns. Especially when captured images are low-resolution…