4 papers
AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models
Andreas Einwiller, Max Klabunde, Florian Lemmerich
The worldwide rise of authoritarianism and the growing role of Large Language Models (LLMs) in users' everyday lives raise the question of whether specific models exhibit or promot…
Revisiting the Relation Between Robustness and Universality
M. Klabunde, L. Caspari, F. Lemmerich
The modified universality hypothesis proposed by Jones et al. (2022) suggests that adversarially robust models trained for a given task are highly similar. We revisit the hypothesi…
Similarity of Neural Network Models: A Survey of Functional and Representational Measures
Max Klabunde, Tobias Schumacher, Markus Strohmaier +1
Measuring similarity of neural networks to understand and improve their behavior has become an issue of great importance and research interest. In this survey, we provide a compreh…
ReSi: A Comprehensive Benchmark for Representational Similarity Measures
Max Klabunde, Tassilo Wald, Tobias Schumacher +3
Measuring the similarity of different representations of neural architectures is a fundamental task and an open research challenge for the machine learning community. This paper pr…