collaborators

15 papers

cs.CL2026

Language Mutations Sustain the Persistences of Conspiracy Theories on Social Media

Calvin Yixiang Cheng, Dorian Quelle, Scott A. Hale

This study investigates how language mutations affect the persistent diffusion of conspiracy theories on social media. Drawing on a three-year dataset of conspiracy-related posts f…

cs.CL2026

PRISM-X: Experiments on Personalised Fine-Tuning with Human and Simulated Users

Hannah Rose Kirk, Liu Leqi, Fanzhi Zeng +4

Personalisation is a standard feature of conversational AI systems used by millions; yet, the efficacy of personalisation methods is often evaluated in academic research using simu…

cs.CY2026

The Enforcement and Feasibility of Hate Speech Moderation on Twitter

Manuel Tonneau, Dylan Thurgood, Diyi Liu +7

Online hate speech is associated with substantial social harms, yet it remains unclear how consistently platforms enforce hate speech policies or whether enforcement is feasible at…

cs.HC2026

Neural steering vectors reveal dose and exposure-dependent impacts of human-AI relationships

Hannah Rose Kirk, Henry Davidson, Ed Saunders +4

Humans are increasingly forming parasocial relationships with AI systems, and modern AI shows an increasing tendency to display social and relationship-seeking behaviour. However,…

cs.CL2025

Lost in translation: using global fact-checks to measure multilingual misinformation prevalence, spread, and evolution

Dorian Quelle, Calvin Cheng, Alexandre Bovet +1

Misinformation and disinformation are growing threats in the digital age, affecting people across languages and borders. However, no research has investigated the prevalence of mul…

cs.CL2025

Measuring what Matters: Construct Validity in Large Language Model Benchmarks

Andrew M. Bean, Ryan Othniel Kearns, Angelika Romanou +39

Evaluating large language models (LLMs) is crucial for both assessing their capabilities and identifying safety or robustness issues prior to deployment. Reliably measuring abstrac…