3 papers
cs.CR2026
AI Security Leaderboard: Methodology, Results and Minimal Standard
Jasper Timm, Lukas Struppek, Ziwei Xu +12
The AI Security Leaderboard is an independent benchmark that ranks the safeguards of frontier AI models from least to most secure. It tests models against the FARAI Minimal Stan…
cs.CV2026
Finding DoRI: Discovery of Retained Images in Diffusion Models
Antoni Kowalczuk, Dominik Hintersdorf, Lukas Struppek +3
Text-to-image diffusion models (DMs) have achieved remarkable success in image generation. However, concerns about data privacy and intellectual property remain due to their potent…
cs.LG2024
Navigating Shortcuts, Spurious Correlations, and Confounders: From Origins via Detection to Mitigation
David Steinmann, Felix Divo, Maurice Kraus +4
Shortcuts, also described as Clever Hans behavior, spurious correlations, or confounders, present a significant challenge in machine learning and AI, critically affecting model gen…