collaborators

23 papers

cs.DL2026

Making a Name for Myself: On Academic Naming Policies and their Impact

A Pranav, Vagrant Gautam, Martin Mundt +6

In academic publishing, names connect scholars to their work. When scholars change their names, including for marriage, academic recognition, or gender transition, they may lose cr…

cs.CL2026

PsychoSafe: Eliciting Psychologically-Informed Refusals in Large Language Models

Gianluca Barmina, Federico Torrielli, Sven Harms +7

Large language models (LLMs) routinely face requests that should be refused, creating a trade-off between helpfulness and harm prevention. However, refusals themselves can be helpf…

cs.CL2026

Epistemic Injustice in Language Models: An Audit of Pretraining Filters and Guardrails

Marco Antonio Stranisci, A Pranav, Rossana Damiano +2

Modern language models rely on pretraining filters to remove undesirable content from training corpora and inference-time guardrails to suppress undesirable outputs during deployme…

cs.CL2026

GRUFF: LLM Pronoun Fidelity, Reasoning, and Biases in German

Fabian Mewes, Anne Lauscher, Vagrant Gautam

Third-person singular pronouns have long been used to study stereotypical biases in language models and to test their abilities to reason about reference. More recently, the interp…

cs.CL2026

Greater accessibility can amplify discrimination in generative AI

Carolin Holtermann, Minh Duc Bui, Kaitlyn Zhou +3

Hundreds of millions of people rely on large language models (LLMs) for education, work, and even healthcare. Yet these models are known to reproduce and amplify social biases pres…

cs.CL2026

Detecting Hallucinations in Authentic LLM-Human Interactions

Yujie Ren, Niklas Gruhlke, Anne Lauscher

As large language models (LLMs) are increasingly applied in sensitive domains such as medicine and law, hallucination detection has become a critical task. Although numerous benchm…