activity
20162026
most citedParaphrasing evades detectors of AI-generated text, but retrieval is an effective defense

91 citations · 357 across the 71 of their papers we have counts for

collaborators
Showing 2024 · cs.CLShow all

8 papers · 2 filters

cs.CL2024

Contextualized Evaluations: Judging Language Model Responses to Underspecified Queries

Chaitanya Malaviya, Joseph Chee Chang, Dan Roth +3

Language model users often issue queries that lack specification, where the context under which a query was issued -- such as the user's identity, the query's intent, and the crite…

cs.CL2024

Localizing and Mitigating Errors in Long-form Question Answering

Rachneet Sachdeva, Yixiao Song, Mohit Iyyer +1

Long-form question answering (LFQA) aims to provide thorough and in-depth answers to complex questions, enhancing comprehension. However, such detailed responses are prone to hallu…

cs.CL2024

Interactive Topic Models with Optimal Transport

Garima Dhanania, Sheshera Mysore, Chau Minh Pham +3

Topic models are widely used to analyze document collections. While they are valuable for discovering latent topics in a corpus when analysts are unfamiliar with the corpus, analys…

cs.CL2024

VERISCORE: Evaluating the factuality of verifiable claims in long-form text generation

Yixiao Song, Yekyung Kim, Mohit Iyyer

Existing metrics for evaluating the factuality of long-form text, such as FACTSCORE (Min et al., 2023) and SAFE (Wei et al., 2024), decompose an input text into "atomic claims" and…

cs.CL2024

Suri: Multi-constraint Instruction Following for Long-form Text Generation

Chau Minh Pham, Simeng Sun, Mohit Iyyer

Existing research on instruction following largely focuses on tasks with simple instructions and short responses. In this work, we explore multi-constraint instruction following fo…

cs.CL2024

CaLMQA: Exploring culturally specific long-form question answering across 23 languages

Shane Arora, Marzena Karpinska, Hung-Ting Chen +3

Despite rising global usage of large language models (LLMs), their ability to generate long-form answers to culturally specific questions remains unexplored in many languages. To f…