11 papers
StabilityBench: Benchmarking Instability in LLMs
Emma Kondrup, Zachary Yang, Anne Imouza +1
AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poorly understood due to strong co…
EASE Configuration Facilitates A Reproducible Science of LLM Social Simulations
Sneheel Sarangi, Maximilian Puelma Touzel, Aurélien Bück-Kaeffer +3
LLMs are increasingly deployed to simulate social interactions, yet many of the existing simulators remain ad hoc and monolithic. This lack of architectural standardization prevent…
The Cookbook: Design Space of LLM-based Social Simulations
Aurélien Bück-Kaeffer, Sneheel Sarangi, Maximilian Puelma Touzel +3
Studies attempting to simulate human behavior with grow in numbers while LLM-only social networks have started appearing outside of controlled settings…
Deepfakes in the 2025 Canadian Election: Prevalence, Partisanship, and Platform Dynamics
Victor Livernoche, Andreea Musulan, Zachary Yang +2
Concerns about AI-generated political content are growing, yet there is limited empirical evidence on how deepfakes actually appear and circulate across social platforms during maj…
OpenFake: An Open Dataset and Platform Toward Real-World Deepfake Detection
Victor Livernoche, Akshatha Arodi, Andreea Musulan +5
Deepfakes, synthetic media created using advanced AI techniques, pose a growing threat to information integrity, particularly in politically sensitive contexts. This challenge is a…
: A Social Media User Dataset for LLM Persona Evaluation and Training
Aurélien Bück-Kaeffer, Je Qin Chooi, Dan Zhao +5
Large language models (LLMs) offer promising capabilities for simulating social media dynamics at scale, enabling studies that would be ethically or logistically challenging with h…