activity
20182022
most citedNTIRE 2020 Challenge on Real-World Image Super-Resolution: Methods and Results

23 citations · 45 across the 8 of their papers we have counts for

collaborators
Showing cs.CVShow all

5 papers · 1 filter

cs.CV20221 cited

A Masked Bounding-Box Selection Based ResNet Predictor for Text Rotation Prediction

Michael Yang, Yuan Lin, ChiuMan Ho

The existing Optical Character Recognition (OCR) systems are capable of recognizing images with horizontal texts. However, when the rotation of the texts increases, it becomes hard…

cs.CV20223 cited

Dual-Flattening Transformers through Decomposed Row and Column Queries for Semantic Segmentation

Ying Wang, Chiuman Ho, Wenju Xu +3

It is critical to obtain high resolution features with long range dependency for dense prediction tasks such as semantic segmentation. To generate high-resolution output of size $H…

cs.CV20212 cited

RSCA: Real-time Segmentation-based Context-Aware Scene Text Detection

Jiachen Li, Yuan Lin, Rongrong Liu +2

Segmentation-based scene text detection methods have been widely adopted for arbitrary-shaped text detection recently, since they make accurate pixel-level predictions on curved te…

cs.CV20211 cited

GCF-Net: Gated Clip Fusion Network for Video Action Recognition

Jenhao Hsiao, Jiawei Chen, Chiuman Ho

In recent years, most of the accuracy gains for video action recognition have come from the newly designed CNN architectures (e.g., 3D-CNNs). These models are trained by applying a…

cs.CV20204 cited

Residual Frames with Efficient Pseudo-3D CNN for Human Action Recognition

Jiawei Chen, Jenson Hsiao, Chiu Man Ho

Human action recognition is regarded as a key cornerstone in domains such as surveillance or video understanding. Despite recent progress in the development of end-to-end solutions…