activity
20192022
most citedVideo Cloze Procedure for Self-Supervised Spatio-Temporal Learning

24 citations · 115 across the 17 of their papers we have counts for

collaborators

22 papers

cs.CV2022

UNITS: Unsupervised Intermediate Training Stage for Scene Text Detection

Youhui Guo, Yu Zhou, Xugong Qin +2

Recent scene text detection methods are almost based on deep learning and data-driven. Synthetic data is commonly adopted for pre-training due to expensive annotation cost. However…

cs.CV20214 cited

Dense Semantic Contrast for Self-Supervised Visual Representation Learning

Xiaoni Li, Yu Zhou, Yifei Zhang +5

Self-supervised representation learning for visual pre-training has achieved remarkable success with sample (instance or pixel) discrimination and semantics discovery of instance,…

cs.CV20211 cited

PIMNet: A Parallel, Iterative and Mimicking Network for Scene Text Recognition

Zhi Qiao, Yu Zhou, Jin Wei +5

Nowadays, scene text recognition has attracted more and more attention due to its various applications. Most state-of-the-art methods adopt an encoder-decoder framework with attent…

cs.CV20211 cited

Which and Where to Focus: A Simple yet Accurate Framework for Arbitrary-Shaped Nearby Text Detection in Scene Images

Youhui Guo, Yu Zhou, Xugong Qin +1

Scene text detection has drawn the close attention of researchers. Though many methods have been proposed for horizontal and oriented texts, previous methods may not perform well w…

cs.CV20214 cited

Mask is All You Need: Rethinking Mask R-CNN for Dense and Arbitrary-Shaped Scene Text Detection

Xugong Qin, Yu Zhou, Youhui Guo +5

Due to the large success in object detection and instance segmentation, Mask R-CNN attracts great attention and is widely adopted as a strong baseline for arbitrary-shaped scene te…

cs.CV202110 cited

Video 3D Sampling for Self-supervised Representation Learning

Wei Li, Dezhao Luo, Bo Fang +2

Most of the existing video self-supervised methods mainly leverage temporal signals of videos, ignoring that the semantics of moving objects and environmental information are all c…