activity
20242026
collaborators

7 papers

cs.CL2026

How Anthropomorphic Language Impacts Public Perceptions of AI

Betty Li Hou, Sophie Hao, Sunoo Park +1

Public discourse about artificial intelligence (AI) often uses anthropomorphic language: language that attributes human capabilities and characteristics to the system. This practic…

cs.LG2026

A Theory of Training Profit-Optimal LLMs

Sophie Hao, William Merrill

Scaling LLMs requires tremendous computational resources, and recent advances in AI have gone hand in hand with massive amounts of capital expenditure. While it is established that…

cs.LG2026

Context-Free Recognition with Transformers

Selim Jerad, Anej Svete, Sophie Hao +2

Transformers excel empirically on tasks that process well-formed inputs according to some grammar, such as natural language and code. However, it remains unclear how they can proce…

cs.CL2025

ModelCitizens: Representing Community Voices in Online Safety

Ashima Suvarna, Christina Chance, Karolina Naranjo +4

Automatic toxic language detection is critical for creating safe, inclusive online spaces. However, it is a highly subjective task, with perceptions of toxic language shaped by com…

cs.CL2025

What Goes Into a LM Acceptability Judgment? Rethinking the Impact of Frequency and Length

Lindia Tjuatja, Graham Neubig, Tal Linzen +1

When comparing the linguistic capabilities of language models (LMs) with humans using LM probabilities, factors such as the length of the sequence and the unigram frequency of lexi…

cs.CL2025

Generative Linguistics, Large Language Models, and the Social Nature of Scientific Success

Sophie Hao

Chesi's (forthcoming) target paper depicts a generative linguistics in crisis, foreboded by Piantadosi's (2023) declaration that "modern language models refute Chomsky's approach t…