82 citations · 157 across the 38 of their papers we have counts for
65 papers · 1 filter
Recovering Diversity Without Losing Alignment: A DPO Recipe for Post-Trained LLMs
Vinay Samuel, Yapei Chang, Mohit Iyyer
Many open-ended instructions have multiple valid answers that users can benefit from seeing, but post-training often narrows an LLM's output space toward a small set of canonical r…
POLARIS: Guiding Small Models to Write Long Stories
Rishanth Rajendhran, Jenna Russell, Mohit Iyyer +1
Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quality significantly degrades as l…
Argument Collapse: LLMs Flatten Long-Form Public Debate
Yekyung Kim, Yapei Chang, Chau Minh Pham +1
As LLMs are increasingly used to draft publicfacing arguments, they may flatten public debate by repeatedly introducing the same polished, plausible arguments. We study argument co…
Beyond Precision: Importance-Aware Recall for Factuality Evaluation in Long-Form LLM Generation
Nazanin Jafari, James Allan, Mohit Iyyer
Evaluating the factuality of long-form output generated by large language models (LLMs) remains challenging, particularly when responses are open-ended and contain many fine-graine…
StoryScope: Investigating idiosyncrasies in AI fiction
Jenna Russell, Rishanth Rajendhran, Chau Minh Pham +2
As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluated. While most existing work in t…
EditLens: Quantifying the Extent of AI Editing in Text
Katherine Thai, Bradley Emi, Elyas Masrour +1
A significant proportion of queries to large language models ask them to edit user-provided text, rather than generate new text from scratch. While previous work focuses on detecti…