51 citations · 91 across the 13 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.CR2023
The Adversarial Implications of Variable-Time Inference
Dudi Biton, Aditi Misra, Efrat Levy +6
Machine learning (ML) models are known to be vulnerable to a number of attacks that target the integrity of their predictions or the privacy of their training data. To carry out th…
cs.CR2023★ 14 cited
Abusing Images and Sounds for Indirect Instruction Injection in Multi-Modal LLMs
Eugene Bagdasaryan, Tsung-Yin Hsieh, Ben Nassi +1
We demonstrate how images and sounds can be used for indirect prompt and instruction injection in multi-modal LLMs. An attacker generates an adversarial perturbation corresponding…