activity
20212024
most citedSystematic Rectification of Language Models via Dead-end Analysis

2 citations · 5 across the 8 of their papers we have counts for

collaborators

8 papers

cs.CL2024

ECBD: Evidence-Centered Benchmark Design for NLP

Yu Lu Liu, Su Lin Blodgett, Jackie Chi Kit Cheung +3

Benchmarking is seen as critical to assessing progress in NLP. However, creating a benchmark involves many design decisions (e.g., which datasets to include, which metrics to use)…

cs.CL2024

A Controlled Reevaluation of Coreference Resolution Models

Ian Porada, Xiyuan Zou, Jackie Chi Kit Cheung

All state-of-the-art coreference resolution (CR) models involve finetuning a pretrained language model. Whether the superior performance of one CR model over another is due to the…

cs.CL20231 cited

Responsible AI Considerations in Text Summarization Research: A Review of Current Practices

Yu Lu Liu, Meng Cao, Su Lin Blodgett +3

AI and NLP publication venues have increasingly encouraged researchers to reflect on possible ethical considerations, adverse impacts, and other responsible AI issues their work mi…

cs.CL2023

Successor Features for Efficient Multisubject Controlled Text Generation

Meng Cao, Mehdi Fatemi, Jackie Chi Kit Cheung +1

While large language models (LLMs) have achieved impressive performance in generating fluent and realistic text, controlling the generated text so that it exhibits properties such…

cs.CL20231 cited

Vārta: A Large-Scale Headline-Generation Dataset for Indic Languages

Rahul Aralikatte, Ziling Cheng, Sumanth Doddapaneni +1

We present Vārta, a large-scale multilingual dataset for headline generation in Indic languages. This dataset includes 41.8 million news articles in 14 different Indic languages (a…

cs.CL20232 cited

Systematic Rectification of Language Models via Dead-end Analysis

Meng Cao, Mehdi Fatemi, Jackie Chi Kit Cheung +1

With adversarial or otherwise normal prompts, existing large language models (LLM) can be pushed to generate toxic discourses. One way to reduce the risk of LLMs generating undesir…