4 papers
GraphKD: Exploring Knowledge Distillation Towards Document Object Detection with Structured Graph Creation
Ayan Banerjee, Sanket Biswas, Josep Lladós +1
Object detection in documents is a key step to automate the structural elements identification process in a digital or scanned document through understanding the hierarchical struc…
Synthetic dataset of ID and Travel Document
Carlos Boned, Maxime Talarmain, Nabil Ghanmi +4
This paper presents a new synthetic dataset of ID and travel documents, called SIDTD. The SIDTD dataset is created to help training and evaluating forged ID documents detection sys…
Harnessing the Power of Multi-Lingual Datasets for Pre-training: Towards Enhancing Text Spotting Performance
Alloy Das, Sanket Biswas, Ayan Banerjee +3
The adaptation capability to a wide range of domains is crucial for scene text spotting models when deployed to real-world conditions. However, existing state-of-the-art (SOTA) app…
Beyond Document Page Classification: Design, Datasets, and Challenges
Jordy Van Landeghem, Sanket Biswas, Matthew B. Blaschko +1
This paper highlights the need to bring document classification benchmarking closer to real-world applications, both in the nature of data tested (: multi-channel, multi-paged,…