5 citations · 6 across the 2 of their papers we have counts for
5 papers
Diacritics Restoration using BERT with Analysis on Czech language
Jakub Náplava, Milan Straka, Jana Straková
We propose a new architecture for diacritics restoration based on contextualized embeddings, namely BERT, and we evaluate it on 12 languages with diacritics. Furthermore, we conduc…
UDPipe at EvaLatin 2020: Contextualized Embeddings and Treebank Embeddings
Milan Straka, Jana Straková
We present our contribution to the EvaLatin shared task, which is the first evaluation campaign devoted to the evaluation of NLP tools for Latin. We submitted a system based on UDP…
Czech Text Processing with Contextual Embeddings: POS Tagging, Lemmatization, Parsing and NER
Milan Straka, Jana Straková, Jan Hajič
Contextualized embeddings, which capture appropriate word meaning depending on context, have recently been proposed. We evaluate two meth ods for precomputing such embeddings, BERT…
Evaluating Contextualized Embeddings on 54 Languages in POS Tagging, Lemmatization and Dependency Parsing
Milan Straka, Jana Straková, Jan Hajič
We present an extensive evaluation of three recently proposed methods for contextualized embeddings on 89 corpora in 54 languages of the Universal Dependencies 2.3 in three tasks:…
UDPipe at SIGMORPHON 2019: Contextualized Embeddings, Regularization with Morphological Categories, Corpora Merging
Milan Straka, Jana Straková, Jan Hajič
We present our contribution to the SIGMORPHON 2019 Shared Task: Crosslinguality and Context in Morphology, Task 2: contextual morphological analysis and lemmatization. We submitted…