3 papers
cs.LG2026
Closing the Distribution Gap in Adversarial Training for LLMs
Chengzhi Hu, Jonas Dornbusch, David Lüdke +2
Adversarial training for LLMs is one of the most promising methods to reliably improve robustness against adversaries. However, despite significant progress, models remain vulnerab…
cs.AI2025
AdversariaLLM: A Unified and Modular Toolbox for LLM Robustness Research
Tim Beyer, Jonas Dornbusch, Jakob Steimle +3
The rapid expansion of research on Large Language Model (LLM) safety and robustness has produced a fragmented and oftentimes buggy ecosystem of implementations, datasets, and evalu…
cs.CV2025
A Simple Combination of Diffusion Models for Better Quality Trade-Offs in Image Denoising
Jonas Dornbusch, Emanuel Pfarr, Florin-Alexandru Vasluianu +2
Diffusion models have garnered considerable interest in computer vision, owing both to their capacity to synthesize photorealistic images and to their proven effectiveness in image…