1 paper
Alessio Chen, Simone Giovannini, Andrea Gemelli +2
Vision-Language Models (VLMs) have shown strong capabilities in document understanding, particularly in identifying and extracting textual information from complex documents. Despi…