3 papers
cs.CR2026
A Sentence Relation-Based Approach to Sanitizing Malicious Instructions
Soumil Datta, Melissa Umble, Daniel S. Brown +1
Retrieval-augmented generation and tool-integrated LLM agents increasingly depend on external textual sources. This reliance broadens the available attack surface, allowing adversa…
cs.LG2026
How Vulnerable Is My Learned Policy? Universal Adversarial Perturbation Attacks On Modern Behavior Cloning Policies
Akansha Kalra, Basavasagar Patil, Guanhong Tao +1
Learning from demonstrations is a popular approach to train AI models; however, their vulnerability to adversarial attacks remains underexplored. We present the first systematic st…
cs.LG2025
Dataset Poisoning Attacks on Behavioral Cloning Policies
Akansha Kalra, Soumil Datta, Ethan Gilmore +3
Behavior Cloning (BC) is a popular framework for training sequential decision policies from expert demonstrations via supervised learning. As these policies are increasingly being…