activity
20242026
collaborators

11 papers

cs.LG2026

StabilityBench: Benchmarking Instability in LLMs

Emma Kondrup, Zachary Yang, Anne Imouza +1

AI Assistants are increasingly deployed in high-stakes settings, such as healthcare or government services. Yet their real-world behavior remains poorly understood due to strong co…

cs.MA2026

EASE Configuration Facilitates A Reproducible Science of LLM Social Simulations

Sneheel Sarangi, Maximilian Puelma Touzel, Aurélien Bück-Kaeffer +3

LLMs are increasingly deployed to simulate social interactions, yet many of the existing simulators remain ad hoc and monolithic. This lack of architectural standardization prevent…

cs.MA2026

The Cookbook: Design Space of LLM-based Social Simulations

Aurélien Bück-Kaeffer, Sneheel Sarangi, Maximilian Puelma Touzel +3

Studies attempting to simulate human behavior with grow in numbers while LLM-only social networks have started appearing outside of controlled settings…

cs.SI2025

Deepfakes in the 2025 Canadian Election: Prevalence, Partisanship, and Platform Dynamics

Victor Livernoche, Andreea Musulan, Zachary Yang +2

Concerns about AI-generated political content are growing, yet there is limited empirical evidence on how deepfakes actually appear and circulate across social platforms during maj…

cs.CV2025

OpenFake: An Open Dataset and Platform Toward Real-World Deepfake Detection

Victor Livernoche, Akshatha Arodi, Andreea Musulan +5

Deepfakes, synthetic media created using advanced AI techniques, pose a growing threat to information integrity, particularly in politically sensitive contexts. This challenge is a…

cs.CL2025

: A Social Media User Dataset for LLM Persona Evaluation and Training

Aurélien Bück-Kaeffer, Je Qin Chooi, Dan Zhao +5

Large language models (LLMs) offer promising capabilities for simulating social media dynamics at scale, enabling studies that would be ethically or logistically challenging with h…