6 citations · 17 across the 24 of their papers we have counts for
4 papers · 1 filter
On Event Individuation for Document-Level Information Extraction
William Gantt, Reno Kriz, Yunmo Chen +2
As information extraction (IE) systems have grown more adept at processing whole documents, the classic task of template filling has seen renewed interest as benchmark for document…
Ambiguous Images With Human Judgments for Robust Visual Event Classification
Kate Sanders, Reno Kriz, Anqi Liu +1
Contemporary vision benchmarks predominantly consider tasks on which humans can achieve near-perfect performance. However, humans are frequently presented with visual data that the…
GEMv2: Multilingual NLG Benchmarking in a Single Line of Code
Sebastian Gehrmann, Abhik Bhattacharjee, Abinaya Mahendiran +74
Evaluation in machine learning is usually informed by past choices, for example which datasets or metrics to use. This standardization enables the comparison on equal footing using…
Creating Multimedia Summaries Using Tweets and Videos
Anietie Andy, Siyi Liu, Daphne Ippolito +3
While popular televised events such as presidential debates or TV shows are airing, people provide commentary on them in real-time. In this paper, we propose a simple yet effective…