4 papers
Checkpoint-GCG: Auditing and Attacking Fine-Tuning-Based Prompt Injection Defenses
Xiaoxue Yang, Bozhidar Stevanoski, Matthieu Meeus +1
Large language models (LLMs) are increasingly deployed in real-world applications ranging from chatbots to agentic systems, where they are expected to process untrusted data and fo…
DeSIA: Attribute Inference Attacks Against Limited Fixed Aggregate Statistics
Yifeng Mao, Bozhidar Stevanoski, Yves-Alexandre de Montjoye
Empirical inference attacks are a popular approach for evaluating the privacy risk of data release mechanisms in practice. While an active attack literature exists to evaluate mach…
Watermarking Training Data of Music Generation Models
Pascal Epple, Igor Shilov, Bozhidar Stevanoski +1
Generative Artificial Intelligence (Gen-AI) models are increasingly used to produce content across domains, including text, images, and audio. While these models represent a major…
QueryCheetah: Fast Automated Discovery of Attribute Inference Attacks Against Query-Based Systems
Bozhidar Stevanoski, Ana-Maria Cretu, Yves-Alexandre de Montjoye
Query-based systems (QBSs) are one of the key approaches for sharing data. QBSs allow analysts to request aggregate information from a private protected dataset. Attacks are a cruc…