8 papers · 1 filter
Introducing corpora Hlava Cor and Hlava AD: Human Label Variation in Coreference and Discourse Relations
Anna Nedoluzhko, Šárka Zikánová, Jiří Mírovský +2
As previous research on annotator disagreement in discourse phenomena has shown, understanding text coherence varies considerably from one individual to another. To explore this ph…
CorPipe at CRAC 2026: Empty Nodes and Cross-Lingual Transfer in Multilingual Coreference Resolution
Milan Straka
We introduce CorPipe 26, our winning submission to the CRAC 2026 Shared Task on Multilingual Coreference Resolution. The fifth edition of this shared task focuses mainly on the com…
Findings of the Fifth Shared Task on Multilingual Coreference Resolution: Expanding Datasets for Long-Range Entities
Michal Novák, Miloslav Konopík, Anna Nedoluzhko +6
This paper describes the fifth edition of the Shared Task on Multilingual Coreference Resolution, held in conjunction with the CODI-CRAC 2026 workshop. Building on previous iterati…
NameTag 3: A Tool and a Service for Multilingual/Multitagset NER
Jana Straková, Milan Straka
We introduce NameTag 3, an open-source tool and cloud-based web service for multilingual, multidataset, and multitagset named entity recognition (NER), supporting both flat and nes…
CorPipe at CRAC 2024: Predicting Zero Mentions from Raw Text
Milan Straka
We present CorPipe 24, the winning entry to the CRAC 2024 Shared Task on Multilingual Coreference Resolution. In this third iteration of the shared task, a novel objective is to al…
Open-Source Web Service with Morphological Dictionary-Supplemented Deep Learning for Morphosyntactic Analysis of Czech
Milan Straka, Jana Straková
We present an open-source web service for Czech morphosyntactic analysis. The system combines a deep learning model with rescoring by a high-precision morphological dictionary at i…