MMOCR: A Comprehensive Toolbox for Text Detection, Recognition and Understanding
arXiv:2108.06543
Abstract
We present MMOCR-an open-source toolbox which provides a comprehensive pipeline for text detection and recognition, as well as their downstream tasks such as named entity recognition and key information extraction. MMOCR implements 14 state-of-the-art algorithms, which is significantly more than all the existing open-source OCR projects we are aware of to date. To facilitate future research and industrial applications of text recognition-related problems, we also provide a large number of trained models and detailed benchmarks to give insights into the performance of text detection, recognition and understanding. MMOCR is publicly released at https://github.com/open-mmlab/mmocr.
Accepted to ACM MM (Open Source Competition Track)
References in corpus (7)
- Rosetta: Large scale system for text detection and recognition in images
- Joint Extraction of Entities and Relations Based on a Novel Tagging Scheme
- Pyramid Mask Text Detector
- CLUENER2020: Fine-grained Named Entity Recognition Dataset and Benchmark for Chinese
- Fourier Contour Embedding for Arbitrary-Shaped Text Detection
- Spatial Dual-Modality Graph Reasoning for Key Information Extraction
- PGNet: Real-time Arbitrarily-Shaped Text Spotting with Point Gathering Network