55 citations · 78 across the 5 of their papers we have counts for
6 papers
UniVIP: A Unified Framework for Self-Supervised Visual Pre-training
Zhaowen Li, Yousong Zhu, Fan Yang +9
Self-supervised learning (SSL) holds promise in leveraging large amounts of unlabeled data. However, the success of popular SSL methods has limited on single-centric-object images…
Recurrent Calibration Network for Irregular Text Recognition
Yunze Gao, Yingying Chen, Jinqiao Wang +3
Scene text recognition has received increased attention in the research community. Text in the wild often possesses irregular arrangements, typically including perspective text, cu…
BDD100K: A Diverse Driving Dataset for Heterogeneous Multitask Learning
Fisher Yu, Haofeng Chen, Xin Wang +5
Datasets drive vision progress, yet existing driving datasets are impoverished in terms of visual content and supported tasks to study multitask learning for autonomous driving. Re…
Reading Scene Text with Attention Convolutional Sequence Modeling
Yunze Gao, Yingying Chen, Jinqiao Wang +1
Reading text in the wild is a challenging task in the field of computer vision. Existing approaches mainly adopted Connectionist Temporal Classification (CTC) or Attention models b…
Fast Deep Matting for Portrait Animation on Mobile Phone
Bingke Zhu, Yingying Chen, Jinqiao Wang +3
Image matting plays an important role in image and video editing. However, the formulation of image matting is inherently ill-posed. Traditional methods usually employ interaction…
Joint Background Reconstruction and Foreground Segmentation via A Two-stage Convolutional Neural Network
Xu Zhao, Yingying Chen, Ming Tang +1
Foreground segmentation in video sequences is a classic topic in computer vision. Due to the lack of semantic and prior knowledge, it is difficult for existing methods to deal with…