1 citations · 1 across the 3 of their papers we have counts for
3 papers
cs.LG2025★ 1 cited
Adversarial Suffix Filtering: a Defense Pipeline for LLMs
David Khachaturov, Robert Mullins
Large Language Models (LLMs) are increasingly embedded in autonomous systems and public-facing environments, yet they remain susceptible to jailbreak vulnerabilities that may under…
cs.LG2025
Watermarking Needs Input Repetition Masking
David Khachaturov, Robert Mullins, Ilia Shumailov +1
Recent advancements in Large Language Models (LLMs) raised concerns over potential misuse, such as for spreading misinformation. In response two counter measures emerged: machine l…
cs.CV2023
Human-Producible Adversarial Examples
David Khachaturov, Yue Gao, Ilia Shumailov +3
Visual adversarial examples have so far been restricted to pixel-level image manipulations in the digital world, or have required sophisticated equipment such as 2D or 3D printers…