1 citations · 1 across the 2 of their papers we have counts for
4 papers
KVP10k : A Comprehensive Dataset for Key-Value Pair Extraction in Business Documents
Oshri Naparstek, Roi Pony, Inbar Shapira +15
In recent years, the challenge of extracting information from business documents has emerged as a critical task, finding applications across numerous domains. This effort has attra…
Optimized Table Tokenization for Table Structure Recognition
Maksym Lysak, Ahmed Nassar, Nikolaos Livathinos +2
Extracting tables from documents is a crucial task in any document conversion pipeline. Recently, transformer-based models have demonstrated that table-structure can be recognized…
TableFormer: Table Structure Understanding with Transformers
Ahmed Nassar, Nikolaos Livathinos, Maksym Lysak +1
Tables organize valuable content in a concise and compact representation. This content is extremely valuable for systems such as search engines, Knowledge Graph's, etc, since they…
Robust PDF Document Conversion Using Recurrent Neural Networks
Nikolaos Livathinos, Cesar Berrospi, Maksym Lysak +7
The number of published PDF documents has increased exponentially in recent decades. There is a growing need to make their rich content discoverable to information retrieval tools.…