11 citations · 20 across the 5 of their papers we have counts for
7 papers
PASTRIE: A Corpus of Prepositions Annotated with Supersense Tags in Reddit International English
Michael Kranzlein, Emma Manning, Siyao Peng +4
We present the Prepositions Annotated with Supersense Tags in Reddit International English ("PASTRIE") corpus, a new dataset containing manually annotated preposition supersenses o…
DisCoDisCo at the DISRPT2021 Shared Task: A System for Discourse Segmentation, Classification, and Connective Detection
Luke Gessler, Shabnam Behzad, Yang Janet Liu +3
This paper describes our submission to the DISRPT2021 Shared Task on Discourse Unit Segmentation, Connective Detection, and Relation Classification. Our system, called DisCoDisCo,…
AMALGUM -- A Free, Balanced, Multilayer English Web Corpus
Luke Gessler, Siyao Peng, Yang Liu +3
We present a freely available, genre-balanced English web corpus totaling 4M tokens and featuring a large number of high-quality automatic annotation layers, including dependency t…
A Corpus of Adpositional Supersenses for Mandarin Chinese
Siyao Peng, Yang Liu, Yilun Zhu +3
Adpositions are frequent markers of semantic relations, but they are highly ambiguous and vary significantly from language to language. Moreover, there is a dearth of annotated cor…
Modeling Long-Range Context for Concurrent Dialogue Acts Recognition
Yue Yu, Siyao Peng, Grace Hui Yang
In dialogues, an utterance is a chain of consecutive sentences produced by one speaker which ranges from a short sentence to a thousand-word post. When studying dialogues at the ut…
All Roads Lead to UD: Converting Stanford and Penn Parses to English Universal Dependencies with Multilayer Annotations
Siyao Peng, Amir Zeldes
We describe and evaluate different approaches to the conversion of gold standard corpus data from Stanford Typed Dependencies (SD) and Penn-style constituent trees to the latest En…