30 citations · 31 across the 3 of their papers we have counts for
3 papers
cs.CL2023
People and Places of Historical Europe: Bootstrapping Annotation Pipeline and a New Corpus of Named Entities in Late Medieval Texts
Vít Novotný, Kristýna Luger, Michal Štefánik +2
Although pre-trained named entity recognition (NER) models are highly accurate on modern corpora, they underperform on historical texts due to differences in language OCR errors. I…
cs.CL2023★ 1 cited
Evaluation of Automatically Constructed Word Meaning Explanations
Marie Stará, Pavel Rychlý, Aleš Horák
Preparing exact and comprehensive word meaning explanations is one of the key steps in the process of monolingual dictionary writing. In standard methodology, the explanations need…
cs.CL2022★ 30 cited
Information Extraction from Scanned Invoice Images using Text Analysis and Layout Features
Hien Thi Ha, Aleš Horák
While storing invoice content as metadata to avoid paper document processing may be the future trend, almost all of daily issued invoices are still printed on paper or generated in…