From the 1 of 7 linked papers with an AI index.
7 papers
Pangram 4 Technical Report
Ben Glickenhaus, Katherine Thai, Jenna Russell +4
The paper introduces Pangram 4, a deep‑learning model for detecting AI‑generated text that achieves high accuracy, strong out‑of‑distribution robustness, and improved detection of…
POLARIS: Guiding Small Models to Write Long Stories
Rishanth Rajendhran, Jenna Russell, Mohit Iyyer +1
Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quality significantly degrades as l…
AI use in American newspapers is widespread, uneven, and rarely disclosed
Jenna Russell, Marzena Karpinska, Destiny Akinode +4
AI is rapidly transforming journalism, but the extent of its use in published newspaper articles remains unclear. We address this gap by auditing a large-scale dataset of 186K arti…
Frankentext: Stitching random text fragments into long-form narratives
Chau Minh Pham, Jenna Russell, Dzung Pham +1
We introduce Frankentexts, a long-form narrative generation paradigm that treats an LLM as a composer of existing texts rather than as an author. Given a writing prompt and thousan…
StoryScope: Investigating idiosyncrasies in AI fiction
Jenna Russell, Rishanth Rajendhran, Chau Minh Pham +2
As AI-generated fiction becomes increasingly prevalent, questions of authorship and originality are becoming central to how written work is evaluated. While most existing work in t…
One ruler to measure them all: Benchmarking multilingual long-context language models
Yekyung Kim, Jenna Russell, Marzena Karpinska +1
We present ONERULER, a multilingual benchmark designed to evaluate long-context language models across 26 languages. ONERULER adapts the English-only RULER benchmark (Hsieh et al.,…