76 citations · 91 across the 3 of their papers we have counts for
9 papers · 1 filter
THEaiTRE 1.0: Interactive generation of theatre play scripts
Rudolf Rosa, Tomáš Musil, Ondřej Dušek +13
We present the first version of a system for interactive generation of theatre play scripts. The system is based on a vanilla GPT-2 model with several adjustments, targeting specif…
Predicting Typological Features in WALS using Language Embeddings and Conditional Probabilities: ÚFAL Submission to the SIGTYP 2020 Shared Task
Martin Vastl, Daniel Zeman, Rudolf Rosa
We present our submission to the SIGTYP 2020 Shared Task on the prediction of typological features. We submit a constrained system, predicting typological features only based on th…
Universal Dependencies according to BERT: both more specific and more general
Tomasz Limisiewicz, Rudolf Rosa, David Mareček
This work focuses on analyzing the form and extent of syntactic abstraction captured by BERT by extracting labeled dependency trees from self-attentions. Previous work showed that…
On the Language Neutrality of Pre-trained Multilingual Representations
Jindřich Libovický, Rudolf Rosa, Alexander Fraser
Multilingual contextual embeddings, such as multilingual BERT and XLM-RoBERTa, have proved useful for many multi-lingual tasks. Previous work probed the cross-linguality of the rep…
How Language-Neutral is Multilingual BERT?
Jindřich Libovický, Rudolf Rosa, Alexander Fraser
Multilingual BERT (mBERT) provides sentence representations for 104 languages, which are useful for many multi-lingual tasks. Previous work probed the cross-linguality of mBERT usi…
Unsupervised Lemmatization as Embeddings-Based Word Clustering
Rudolf Rosa, Zdeněk Žabokrtský
We focus on the task of unsupervised lemmatization, i.e. grouping together inflected forms of one word under one label (a lemma) without the use of annotated training data. We prop…