activity
20242026
collaborators

5 papers

cs.CL2026

Analysing Differences in Persuasive Language in LLM-Generated Text: Uncovering Stereotypical Gender Patterns

Amalie Brogaard Pauli, Maria Barrett, Max Müller-Eberstein +2

Large language models (LLMs) are increasingly used for everyday communication tasks, including drafting interpersonal messages intended to influence and persuade. Prior work has sh…

cs.CL2026

When Meanings Meet: Investigating the Emergence and Quality of Shared Concept Spaces during Multilingual Language Model Training

Felicia Körner, Max Müller-Eberstein, Anna Korhonen +1

Training Large Language Models (LLMs) with high multilingual coverage is becoming increasingly important -- especially when monolingual resources are scarce. Recent studies have fo…

cs.CL2025

PolyPythias: Stability and Outliers across Fifty Language Model Pre-Training Runs

Oskar van der Wal, Pietro Lesci, Max Muller-Eberstein +4

The stability of language model pre-training and its effects on downstream performance are still understudied. Prior work shows that the training process can yield significantly di…

cs.CL2025

DaKultur: Evaluating the Cultural Awareness of Language Models for Danish with Native Speakers

Max Müller-Eberstein, Mike Zhang, Elisa Bassignana +2

Large Language Models (LLMs) have seen widespread societal adoption. However, while they are able to interact with users in languages beyond English, they have been shown to lack c…

cs.CL2024

SnakModel: Lessons Learned from Training an Open Danish Large Language Model

Mike Zhang, Max Müller-Eberstein, Elisa Bassignana +1

We present SnakModel, a Danish large language model (LLM) based on Llama2-7B, which we continuously pre-train on 13.6B Danish words, and further tune on 3.7M Danish instructions. A…