1 paper
Khanh Nguyen, Raouf Kerkouche, Mario Fritz +1
Document Visual Question Answering (DocVQA) has introduced a new paradigm for end-to-end document understanding, and quickly became one of the standard benchmarks for multimodal LL…