2 papers
cs.CV2025
Improving OCR using internal document redundancy
Diego Belzarena, Seginus Mowlavi, Aitor Artola +9
Current OCR systems are based on deep learning models trained on large amounts of data. Although they have shown some ability to generalize to unseen data, especially in detection…
cs.CV2025★ 3 cited
Normalized vs Diplomatic Annotation: A Case Study of Automatic Information Extraction from Handwritten Uruguayan Birth Certificates
Natalia Bottaioli, Solène Tarride, Jérémy Anger +8
This study evaluates the recently proposed Document Attention Network (DAN) for extracting key-value information from Uruguayan birth certificates, handwritten in Spanish. We inves…