24 citations · 115 across the 17 of their papers we have counts for
20 papers
UNITS: Unsupervised Intermediate Training Stage for Scene Text Detection
Youhui Guo, Yu Zhou, Xugong Qin +2
Recent scene text detection methods are almost based on deep learning and data-driven. Synthetic data is commonly adopted for pre-training due to expensive annotation cost. However…
Dense Semantic Contrast for Self-Supervised Visual Representation Learning
Xiaoni Li, Yu Zhou, Yifei Zhang +5
Self-supervised representation learning for visual pre-training has achieved remarkable success with sample (instance or pixel) discrimination and semantics discovery of instance,…
PIMNet: A Parallel, Iterative and Mimicking Network for Scene Text Recognition
Zhi Qiao, Yu Zhou, Jin Wei +5
Nowadays, scene text recognition has attracted more and more attention due to its various applications. Most state-of-the-art methods adopt an encoder-decoder framework with attent…
Which and Where to Focus: A Simple yet Accurate Framework for Arbitrary-Shaped Nearby Text Detection in Scene Images
Youhui Guo, Yu Zhou, Xugong Qin +1
Scene text detection has drawn the close attention of researchers. Though many methods have been proposed for horizontal and oriented texts, previous methods may not perform well w…
Mask is All You Need: Rethinking Mask R-CNN for Dense and Arbitrary-Shaped Scene Text Detection
Xugong Qin, Yu Zhou, Youhui Guo +5
Due to the large success in object detection and instance segmentation, Mask R-CNN attracts great attention and is widely adopted as a strong baseline for arbitrary-shaped scene te…
Video 3D Sampling for Self-supervised Representation Learning
Wei Li, Dezhao Luo, Bo Fang +2
Most of the existing video self-supervised methods mainly leverage temporal signals of videos, ignoring that the semantics of moving objects and environmental information are all c…