activity
20242026
collaborators

5 papers

cs.CV2026

When Low CER is Not Enough: An Analysis of Hallucinations in Vision-Language OCR Systems on Historical Uruguayan Documents

Marina Gardella, Camilo Mari{ñ}o, Diego Belzarena +3

Optical Character Recognition (OCR) is a key component in the digitization of historical archives. Recently, Vision-Language Models (VLMs) have emerged as strong alternatives to tr…

cs.CV2026

CHROMA: Detecting AI-Generated Images through Inter-Channel Color-Space Correlations

Juan Pablo Sotelo, Marina Gardella, Pablo Musé

The rapid adoption of diffusion and large-scale generative models has made it increasingly challenging to distinguish synthetic imagery from real photographs. While automated detec…

cs.CV2025

Improving OCR using internal document redundancy

Diego Belzarena, Seginus Mowlavi, Aitor Artola +9

Current OCR systems are based on deep learning models trained on large amounts of data. Although they have shown some ability to generalize to unseen data, especially in detection…

cs.CV2025

Normalized vs Diplomatic Annotation: A Case Study of Automatic Information Extraction from Handwritten Uruguayan Birth Certificates

Natalia Bottaioli, Solène Tarride, Jérémy Anger +8

This study evaluates the recently proposed Document Attention Network (DAN) for extracting key-value information from Uruguayan birth certificates, handwritten in Spanish. We inves…

cs.CV2024

PhotoHolmes: a Python library for forgery detection in digital images

Julián O'Flaherty, Rodrigo Paganini, Juan Pablo Sotelo +4

In this paper, we introduce PhotoHolmes, an open-source Python library designed to easily run and benchmark forgery detection methods on digital images. The library includes implem…