2 papers
cs.CV2025
Improving OCR using internal document redundancy
Diego Belzarena, Seginus Mowlavi, Aitor Artola +9
Current OCR systems are based on deep learning models trained on large amounts of data. Although they have shown some ability to generalize to unseen data, especially in detection…
cs.CV2025
Normalized vs Diplomatic Annotation: A Case Study of Automatic Information Extraction from Handwritten Uruguayan Birth Certificates
Natalia Bottaioli, Solène Tarride, Jérémy Anger +8
This study evaluates the recently proposed Document Attention Network (DAN) for extracting key-value information from Uruguayan birth certificates, handwritten in Spanish. We inves…