4 papers
Between Help and Harm: An Evaluation of Mental Health Crisis Handling by LLMs
Adrian Arnaiz-Rodriguez, Miguel Baidal, Erik Derner +5
Large language model-powered chatbots have transformed how people seek information, especially in high-stakes contexts like mental health. Despite their support capabilities, safe…
Mind the Style: Impact of Communication Style on Human-Chatbot Interaction
Erik Derner, Dalibor KuÄera, Dalibor Kučera +3
Conversational agents increasingly mediate everyday digital interactions, yet the effects of their communication style on user experience and task success remain insufficiently und…
Large Reasoning Models Are Autonomous Jailbreak Agents
Thilo Hagendorff, Erik Derner, Nuria Oliver
Jailbreaking -- bypassing built-in safety mechanisms in AI models -- has traditionally required complex technical procedures or specialized human expertise. In this study, we show…
Leveraging Large Language Models to Measure Gender Representation Bias in Gendered Language Corpora
Erik Derner, Sara Sansalvador de la Fuente, Yoan Gutiérrez +2
Large language models (LLMs) often inherit and amplify social biases embedded in their training data. A prominent social bias is gender bias. In this regard, prior work has mainly…