activity
20152022
most citedBig Data Small Data, In Domain Out-of Domain, Known Word Unknown Word: The Impact of Word Representation on Sequence Labelling Tasks

23 citations · 28 across the 14 of their papers we have counts for

collaborators
Showing cs.CLShow all

24 papers · 1 filter

cs.CL2022

Measuring Fine-Grained Semantic Equivalence with Abstract Meaning Representation

Shira Wein, Zhuxin Wang, Nathan Schneider

Identifying semantically equivalent sentences is important for many cross-lingual and mono-lingual NLP tasks. Current approaches to semantic equivalence take a loose, sentence-leve…

cs.CL20221 cited

CGELBank: CGEL as a Framework for English Syntax Annotation

Brett Reynolds, Aryaman Arora, Nathan Schneider

We introduce the syntactic formalism of the \textit{Cambridge Grammar of the English Language} (CGEL) to the world of treebanking through the CGELBank project. We discuss some issu…

cs.CL2022

MASALA: Modelling and Analysing the Semantics of Adpositions in Linguistic Annotation of Hindi

Aryaman Arora, Nitin Venkateswaran, Nathan Schneider

We present a completed, publicly available corpus of annotated semantic relations of adpositions and case markers in Hindi. We used the multilingual SNACS annotation scheme, which…

cs.CL2022

Spanish Abstract Meaning Representation: Annotation of a General Corpus

Shira Wein, Lucia Donatelli, Ethan Ricker +3

The Abstract Meaning Representation (AMR) formalism, designed originally for English, has been adapted to a number of languages. We build on previous work proposing the annotation…

cs.CL2021

PASTRIE: A Corpus of Prepositions Annotated with Supersense Tags in Reddit International English

Michael Kranzlein, Emma Manning, Siyao Peng +4

We present the Prepositions Annotated with Supersense Tags in Reddit International English ("PASTRIE") corpus, a new dataset containing manually annotated preposition supersenses o…

cs.CL2021

BERT Has Uncommon Sense: Similarity Ranking for Word Sense BERTology

Luke Gessler, Nathan Schneider

An important question concerning contextualized word embedding (CWE) models like BERT is how well they can represent different word senses, especially those in the long tail of unc…