2 papers
cs.CL2026
SimBench: Benchmarking the Ability of Large Language Models to Simulate Human Behaviors
Tiancheng Hu, Joachim Baumann, Lorenzo Lupo +3
Large language model (LLM) simulations of human behavior have the potential to revolutionize the social and behavioral sciences, if and only if they faithfully reflect real human b…
cs.SI2026
The Content Moderator's Dilemma: Removal of Toxic Content and Distortions to Online Discourse
Mahyar Habibi, Dirk Hovy, Carlo Schwarz
There is an ongoing debate about how to moderate toxic speech on social media and the impact of content moderation on online discourse. This paper proposes and validates a methodol…