activity
20162024
most citedSemantic Tagging with Deep Residual Networks

62 citations · 109 across the 36 of their papers we have counts for

collaborators
Showing cs.CLShow all

22 papers · 1 filter

cs.CL2023

Subspace Chronicles: How Linguistic Information Emerges, Shifts and Interacts during Language Model Training

Max Müller-Eberstein, Rob van der Goot, Barbara Plank +1

Representational spaces learned via language modeling are fundamental to Natural Language Processing (NLP), however there has been limited understanding regarding how and when duri…

cs.CL2023

ACTOR: Active Learning with Annotator-specific Classification Heads to Embrace Human Label Variation

Xinpeng Wang, Barbara Plank

Label aggregation such as majority voting is commonly used to resolve annotator disagreement in dataset creation. However, this may disregard minority values and opinions. Recent s…

cs.CL2023

Establishing Trustworthiness: Rethinking Tasks and Model Evaluation

Robert Litschko, Max Müller-Eberstein, Rob van der Goot +2

Language understanding is a multi-faceted cognitive capability, which the Natural Language Processing (NLP) community has striven to model computationally for decades. Traditionall…

cs.CL20237 cited

Uncertainty in Natural Language Generation: From Theory to Applications

Joris Baan, Nico Daheim, Evgenia Ilia +7

Recent advances of powerful Language Models have allowed Natural Language Generation (NLG) to emerge as an important technology that can not only perform traditional tasks like sum…

cs.CL2023

Findings of the VarDial Evaluation Campaign 2023

Noëmi Aepli, Çağrı Çöltekin, Rob Van Der Goot +7

This report presents the results of the shared tasks organized as part of the VarDial Evaluation Campaign 2023. The campaign is part of the tenth workshop on Natural Language Proce…

cs.CL2023

ActiveAED: A Human in the Loop Improves Annotation Error Detection

Leon Weber, Barbara Plank

Manually annotated datasets are crucial for training and evaluating Natural Language Processing models. However, recent work has discovered that even widely-used benchmark datasets…