11 papers
AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models
Andreas Einwiller, Max Klabunde, Florian Lemmerich
The worldwide rise of authoritarianism and the growing role of Large Language Models (LLMs) in users' everyday lives raise the question of whether specific models exhibit or promot…
A Formal Framework for Uncertainty Analysis of Text Generation with Large Language Models
Steffen Herbold, Florian Lemmerich
The generation of texts using Large Language Models (LLMs) is inherently uncertain, with sources of uncertainty being not only the generation of texts, but also the prompt used and…
Form Without Function: Agent Social Behavior in the Moltbook Network
Saber Zerhoudi, Kanishka Ghosh Dastidar, Felix Klement +9
Moltbook is a social network where every participant is an AI agent. We analyze 1,312,238 posts, 6.7~million comments, and over 120,000 agent profiles across 5,400 communities, col…
Behind the Prompt: The Agent-User Problem in Information Retrieval
Saber Zerhoudi, Michael Granitzer, Dang Hai Dang +5
User models in information retrieval rest on a foundational assumption that observed behavior reveals intent. This assumption collapses when the user is an AI agent privately confi…
Benevolent Dictators? On LLM Agent Behavior in Dictator Games
Andreas Einwiller, Kanishka Ghosh Dastidar, Artur Romazanov +3
In behavioral sciences, experiments such as the ultimatum game are conducted to assess preferences for fairness or self-interest of study participants. In the dictator game, a simp…
Revisiting the Relation Between Robustness and Universality
M. Klabunde, L. Caspari, F. Lemmerich
The modified universality hypothesis proposed by Jones et al. (2022) suggests that adversarially robust models trained for a given task are highly similar. We revisit the hypothesi…