3 papers
cs.HC2026
Towards Non-Latin Text and Layout Personalization for Enhanced Readability
Rina Buoy, Dylan berkamp Fouepe Dongmo, Vesal Khean +2
Reading has always been an integral part of both professional and personal life. Character and layout recognition and understanding by computers are well-explored areas. Neverthele…
cs.CL2025
BoundingDocs: a Unified Dataset for Document Question Answering with Spatial Annotations
Simone Giovannini, Fabio Coppini, Andrea Gemelli +1
We present a unified dataset for document Question-Answering (QA), which is obtained combining several public datasets related to Document AI and visually rich document understandi…
cs.CL2025
Towards Reliable and Interpretable Document Question Answering via VLMs
Alessio Chen, Simone Giovannini, Andrea Gemelli +2
Vision-Language Models (VLMs) have shown strong capabilities in document understanding, particularly in identifying and extracting textual information from complex documents. Despi…