activity
20182023
most citedThe GEM Benchmark: Natural Language Generation, its Evaluation and Metrics

52 citations · 55 across the 6 of their papers we have counts for

collaborators

14 papers

cs.CL20222 cited

Missing Counter-Evidence Renders NLP Fact-Checking Unrealistic for Misinformation

Max Glockner, Yufang Hou, Iryna Gurevych

Misinformation emerges in times of uncertainty when credible information is limited. This is challenging for NLP-based fact-checking as it relies on counter-evidence, which may not…

cs.CL2021

Overview of the 2021 Key Point Analysis Shared Task

Roni Friedman, Lena Dankin, Yufang Hou +3

We describe the 2021 Key Point Analysis (KPA-2021) shared task on key point analysis that we organized as a part of the 8th Workshop on Argument Mining (ArgMining 2021) at EMNLP 20…

cs.CL2021

End-to-end Neural Information Status Classification

Yufang Hou

Most previous studies on information status (IS) classification and bridging anaphora recognition assume that the gold mention or syntactic tree information is given (Hou et al., 2…

cs.CL2021

D2S: Document-to-Slide Generation Via Query-Based Text Summarization

Edward Sun, Yufang Hou, Dakuo Wang +2

Presentations are critical for communication in all areas of our lives, yet the creation of slide decks is often tedious and time-consuming. There has been limited research aiming…

cs.CL2021

Probing for Bridging Inference in Transformer Language Models

Onkar Pandit, Yufang Hou

We probe pre-trained transformer language models for bridging inference. We first investigate individual attention heads in BERT and observe that attention heads at higher layers p…

cs.CL202152 cited

The GEM Benchmark: Natural Language Generation, its Evaluation and Metrics

Sebastian Gehrmann, Tosin Adewumi, Karmanya Aggarwal +53

We introduce GEM, a living benchmark for natural language Generation (NLG), its Evaluation, and Metrics. Measuring progress in NLG relies on a constantly evolving ecosystem of auto…