23 citations · 45 across the 8 of their papers we have counts for
5 papers · 1 filter
A Masked Bounding-Box Selection Based ResNet Predictor for Text Rotation Prediction
Michael Yang, Yuan Lin, ChiuMan Ho
The existing Optical Character Recognition (OCR) systems are capable of recognizing images with horizontal texts. However, when the rotation of the texts increases, it becomes hard…
Dual-Flattening Transformers through Decomposed Row and Column Queries for Semantic Segmentation
Ying Wang, Chiuman Ho, Wenju Xu +3
It is critical to obtain high resolution features with long range dependency for dense prediction tasks such as semantic segmentation. To generate high-resolution output of size $H…
RSCA: Real-time Segmentation-based Context-Aware Scene Text Detection
Jiachen Li, Yuan Lin, Rongrong Liu +2
Segmentation-based scene text detection methods have been widely adopted for arbitrary-shaped text detection recently, since they make accurate pixel-level predictions on curved te…
GCF-Net: Gated Clip Fusion Network for Video Action Recognition
Jenhao Hsiao, Jiawei Chen, Chiuman Ho
In recent years, most of the accuracy gains for video action recognition have come from the newly designed CNN architectures (e.g., 3D-CNNs). These models are trained by applying a…
Residual Frames with Efficient Pseudo-3D CNN for Human Action Recognition
Jiawei Chen, Jenson Hsiao, Chiu Man Ho
Human action recognition is regarded as a key cornerstone in domains such as surveillance or video understanding. Despite recent progress in the development of end-to-end solutions…