ICDAR2019 Competition on Scanned Receipt OCR and Information Extraction
arXiv:2103.10213 · doi:10.1109/ICDAR.2019.00244
Abstract
Scanned receipts OCR and key information extraction (SROIE) represent the processeses of recognizing text from scanned receipts and extracting key texts from them and save the extracted tests to structured documents. SROIE plays critical roles for many document analysis applications and holds great commercial potentials, but very little research works and advances have been published in this area. In recognition of the technical challenges, importance and huge commercial potentials of SROIE, we organized the ICDAR 2019 competition on SROIE. In this competition, we set up three tasks, namely, Scanned Receipt Text Localisation (Task 1), Scanned Receipt OCR (Task 2) and Key Information Extraction from Scanned Receipts (Task 3). A new dataset with 1000 whole scanned receipt images and annotations is created for the competition. In this report we will presents the motivation, competition datasets, task definition, evaluation protocol, submission statistics, performance of submitted methods and results analysis.
Cited by in corpus (18)
- Information Extraction from Scanned Invoice Images using Text Analysis and Layout Features
- A Survey of Deep Learning Approaches for OCR and Document Understanding
- Text Detection and Recognition in the Wild: A Review
- PP-OCRv2: Bag of Tricks for Ultra Lightweight OCR System
- Learning Graph Normalization for Graph Neural Networks
- Abstractive Information Extraction from Scanned Invoices (AIESI) using End-to-end Sequential Approach
- Asking questions on handwritten document collections
- Improving Information Extraction on Business Documents with Specific Pre-Training Tasks
- StrucTexT: Structured Text Understanding with Multi-Modal Transformers
- MatchVIE: Exploiting Match Relevancy between Entities for Visual Information Extraction
- Long-Range Transformer Architectures for Document Understanding
- Key Information Extraction From Documents: Evaluation And Generator
- ViBERTgrid: A Jointly Trained Multi-Modal 2D Document Representation for Key Information Extraction from Documents
- Extracting Complex Named Entities in Legal Documents via Weakly Supervised Object Detection
- Tag, Copy or Predict: A Unified Weakly-Supervised Learning Framework for Visual Information Extraction using Sequences
- Robustness Evaluation of Transformer-based Form Field Extractors via Form Attacks
- Entity Relation Extraction as Dependency Parsing in Visually Rich Documents
- Understanding Scanned Receipts