activity
20182022
most citedTreebanking User-Generated Content: a UD Based Overview of Guidelines, Corpora and Unified Recommendations

12 citations · 32 across the 10 of their papers we have counts for

collaborators

15 papers

cs.CL20224 cited

GCDT: A Chinese RST Treebank for Multigenre and Multilingual Discourse Parsing

Siyao Peng, Yang Janet Liu, Amir Zeldes

A lack of large-scale human-annotated data has hampered the hierarchical discourse parsing of Chinese. In this paper, we present GCDT, the largest hierarchical discourse treebank f…

cs.CL2022

A Second Wave of UD Hebrew Treebanking and Cross-Domain Parsing

Amir Zeldes, Nick Howell, Noam Ordan +1

Foundational Hebrew NLP tasks such as segmentation, tagging and parsing, have relied to date on various versions of the Hebrew Treebank (HTB, Sima'an et al. 2001). However, the dat…

cs.CL2021

Anatomy of OntoGUM--Adapting GUM to the OntoNotes Scheme to Evaluate Robustness of SOTA Coreference Algorithms

Yilun Zhu, Sameer Pradhan, Amir Zeldes

SOTA coreference resolution produces increasingly impressive scores on the OntoNotes benchmark. However lack of comparable data following the same scheme for more genres makes it d…

cs.CL2021

DisCoDisCo at the DISRPT2021 Shared Task: A System for Discourse Segmentation, Classification, and Connective Detection

Luke Gessler, Shabnam Behzad, Yang Janet Liu +3

This paper describes our submission to the DISRPT2021 Shared Task on Discourse Unit Segmentation, Connective Detection, and Relation Classification. Our system, called DisCoDisCo,…

cs.CL2021

WikiGUM: Exhaustive Entity Linking for Wikification in 12 Genres

Jessica Lin, Amir Zeldes

Previous work on Entity Linking has focused on resources targeting non-nested proper named entity mentions, often in data from Wikipedia, i.e. Wikification. In this paper, we prese…

cs.CL20211 cited

OntoGUM: Evaluating Contextualized SOTA Coreference Resolution on 12 More Genres

Yilun Zhu, Sameer Pradhan, Amir Zeldes

SOTA coreference resolution produces increasingly impressive scores on the OntoNotes benchmark. However lack of comparable data following the same scheme for more genres makes it d…