11 citations · 27 across the 3 of their papers we have counts for
11 papers
Prague Dependency Treebank -- Consolidated 1.0
Jan Hajič, Eduard Bejček, Jaroslava Hlaváčová +4
We present a richly annotated and genre-diversified language resource, the Prague Dependency Treebank-Consolidated 1.0 (PDT-C 1.0), the purpose of which is - as it always been the…
European Language Grid: An Overview
Georg Rehm, Maria Berger, Ela Elsholz +33
With 24 official EU and many additional languages, multilingualism in Europe and an inclusive Digital Single Market can only be enabled through Language Technologies (LTs). Europea…
Czech Text Processing with Contextual Embeddings: POS Tagging, Lemmatization, Parsing and NER
Milan Straka, Jana Straková, Jan Hajič
Contextualized embeddings, which capture appropriate word meaning depending on context, have recently been proposed. We evaluate two meth ods for precomputing such embeddings, BERT…
Evaluating Contextualized Embeddings on 54 Languages in POS Tagging, Lemmatization and Dependency Parsing
Milan Straka, Jana Straková, Jan Hajič
We present an extensive evaluation of three recently proposed methods for contextualized embeddings on 89 corpora in 54 languages of the Universal Dependencies 2.3 in three tasks:…
UDPipe at SIGMORPHON 2019: Contextualized Embeddings, Regularization with Morphological Categories, Corpora Merging
Milan Straka, Jana Straková, Jan Hajič
We present our contribution to the SIGMORPHON 2019 Shared Task: Crosslinguality and Context in Morphology, Task 2: contextual morphological analysis and lemmatization. We submitted…
Neural Architectures for Nested NER through Linearization
Jana Straková, Milan Straka, Jan Hajič
We propose two neural network architectures for nested named entity recognition (NER), a setting in which named entities may overlap and also be labeled with more than one label. W…