90 citations · 159 across the 22 of their papers we have counts for
3 papers · 1 filter
AuAu: A Benchmark for Auditing Authoritarian Alignment in Large Language Models
Andreas Einwiller, Max Klabunde, Florian Lemmerich
The worldwide rise of authoritarianism and the growing role of Large Language Models (LLMs) in users' everyday lives raise the question of whether specific models exhibit or promot…
Joint Multiclass Debiasing of Word Embeddings
Radomir Popović, Florian Lemmerich, Markus Strohmaier
Bias in Word Embeddings has been a subject of recent interest, along with efforts for its reduction. Current approaches show promising progress towards debiasing single bias dimens…
The POLAR Framework: Polar Opposites Enable Interpretability of Pre-Trained Word Embeddings
Binny Mathew, Sandipan Sikdar, Florian Lemmerich +1
We introduce POLAR - a framework that adds interpretability to pre-trained word embeddings via the adoption of semantic differentials. Semantic differentials are a psychometric con…