3 papers
cs.AI2026
When Do Data-Driven Systems Exhibit the Capability to Infer?
Maximilian Poretschkin, Tabea Naeven
The European AI Act is the first comprehensive regulation of artificial intelligence (AI), setting out extensive obligations, particularly for so-called high-risk and general-purpo…
cs.CL2025
Textual Data Bias Detection and Mitigation -- An Extensible Pipeline with Experimental Evaluation
Rebekka Görge, Sujan Sai Gannamaneni, Tabea Naeven +10
Textual data used to train large language models (LLMs) exhibits multifaceted bias manifestations encompassing harmful language and skewed demographic distributions. Regulations su…
cs.CL2024
LLMs and Memorization: On Quality and Specificity of Copyright Compliance
Felix B Mueller, Rebekka Görge, Anna K Bernzen +2
Memorization in large language models (LLMs) is a growing concern. LLMs have been shown to easily reproduce parts of their training data, including copyrighted work. This is an imp…