UniMorph 2.0: Universal Morphology
arXiv:1810.11101
Abstract
The Universal Morphology UniMorph project is a collaborative effort to improve how NLP handles complex morphology across the world's languages. The project releases annotated morphological data using a universal tagset, the UniMorph schema. Each inflected form is associated with a lemma, which typically carries its underlying lexical meaning, and a bundle of morphological features from our schema. Additional supporting data and tools are also released on a per-language basis when available. UniMorph is based at the Center for Language and Speech Processing (CLSP) at Johns Hopkins University in Baltimore, Maryland and is sponsored by the DARPA LORELEI program. This paper details advances made to the collection, annotation, and dissemination of project resources since the initial UniMorph release described at LREC 2016. lexical resources} }
LREC 2018
Cited by in corpus (12)
- The CoNLL--SIGMORPHON 2018 Shared Task: Universal Morphological Reinflection
- Morphological Tagging and Lemmatization of Albanian: A Manually Annotated Corpus and Neural Models
- Modelling Verbal Morphology in Nen
- Morphological Irregularity Correlates with Frequency
- Finding the way from ä to a: Sub-character morphological inflection for the SIGMORPHON 2018 Shared Task
- Are All Languages Equally Hard to Language-Model?
- Cross-Lingual Transfer of Semantic Roles: From Raw Text to Semantic Roles
- A Formal Description of Sorani Kurdish Morphology
- Lost in Evaluation: Misleading Benchmarks for Bilingual Dictionary Induction
- Neural Transductive Learning and Beyond: Morphological Generation in the Minimal-Resource Setting
- Unsupervised Disambiguation of Syncretism in Inflected Lexicons
- RuSentEval: Linguistic Source, Encoder Force!