activity
20182021
most citedTEINet: Towards an Efficient Architecture for Video Recognition

28 citations · 33 across the 5 of their papers we have counts for

collaborators
Showing cs.CVShow all

10 papers · 1 filter

cs.CV20212 cited

ARTS: Eliminating Inconsistency between Text Detection and Recognition with Auto-Rectification Text Spotter

Humen Zhong, Jun Tang, Wenhai Wang +3

Recent approaches for end-to-end text spotting have achieved promising results. However, most of the current spotters were plagued by the inconsistency problem between text detecti…

cs.CV2021

PAN++: Towards Efficient and Accurate End-to-End Spotting of Arbitrarily-Shaped Text

Wenhai Wang, Enze Xie, Xiang Li +5

Scene text detection and recognition have been well explored in the past few years. Despite the progress, efficient and accurate end-to-end spotting of arbitrarily-shaped text rema…

cs.CV2021

Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions

Wenhai Wang, Enze Xie, Xiang Li +6

Although using convolutional neural networks (CNNs) as backbones achieves great successes in computer vision, this work investigates a simple backbone network useful for many dense…

cs.CV2020

Frequency Consistent Adaptation for Real World Super Resolution

Xiaozhong Ji, Guangpin Tao, Yun Cao +5

Recent deep-learning based Super-Resolution (SR) methods have achieved remarkable performance on images with known degradation. However, these methods always fail in real-world sce…

cs.CV20203 cited

A New Unified Method for Detecting Text from Marathon Runners and Sports Players in Video

Sauradip Nag, Palaiahnakote Shivakumara, Umapada Pal +2

Detecting text located on the torsos of marathon runners and sports players in video is a challenging issue due to poor quality and adverse effects caused by flexible/colorful clot…

cs.CV201928 cited

TEINet: Towards an Efficient Architecture for Video Recognition

Zhaoyang Liu, Donghao Luo, Yabiao Wang +6

Efficiency is an important issue in designing video architectures for action recognition. 3D CNNs have witnessed remarkable progress in action recognition from videos. However, com…