2 citations · 5 across the 8 of their papers we have counts for
8 papers
ECBD: Evidence-Centered Benchmark Design for NLP
Yu Lu Liu, Su Lin Blodgett, Jackie Chi Kit Cheung +3
Benchmarking is seen as critical to assessing progress in NLP. However, creating a benchmark involves many design decisions (e.g., which datasets to include, which metrics to use)…
A Controlled Reevaluation of Coreference Resolution Models
Ian Porada, Xiyuan Zou, Jackie Chi Kit Cheung
All state-of-the-art coreference resolution (CR) models involve finetuning a pretrained language model. Whether the superior performance of one CR model over another is due to the…
Responsible AI Considerations in Text Summarization Research: A Review of Current Practices
Yu Lu Liu, Meng Cao, Su Lin Blodgett +3
AI and NLP publication venues have increasingly encouraged researchers to reflect on possible ethical considerations, adverse impacts, and other responsible AI issues their work mi…
Successor Features for Efficient Multisubject Controlled Text Generation
Meng Cao, Mehdi Fatemi, Jackie Chi Kit Cheung +1
While large language models (LLMs) have achieved impressive performance in generating fluent and realistic text, controlling the generated text so that it exhibits properties such…
Vārta: A Large-Scale Headline-Generation Dataset for Indic Languages
Rahul Aralikatte, Ziling Cheng, Sumanth Doddapaneni +1
We present Vārta, a large-scale multilingual dataset for headline generation in Indic languages. This dataset includes 41.8 million news articles in 14 different Indic languages (a…
Systematic Rectification of Language Models via Dead-end Analysis
Meng Cao, Mehdi Fatemi, Jackie Chi Kit Cheung +1
With adversarial or otherwise normal prompts, existing large language models (LLM) can be pushed to generate toxic discourses. One way to reduce the risk of LLMs generating undesir…