3 papers
cs.CL2026
Post-training makes large language models less human-like
Marcel Binz, Elif Akata, Abdullah Almaatouq +76
Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavior and why. To address this, w…
cs.CY2026
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
Michelle Vaccaro, Jaeyoon Song, Abdullah Almaatouq +1
Current frontier AI safety evaluations emphasize static benchmarks, third-party annotations, and red-teaming. In this position paper, we argue that AI safety research should focus…
econ.GN2025
Integrative Experiments Identify How Punishment Impacts Welfare in Public Goods Games
Mohammed Alsobay, David G. Rand, Duncan J. Watts +1
Punishment as a mechanism for promoting cooperation has been studied extensively for more than two decades, but its effectiveness remains a matter of dispute. Here, we examine how…