2 papers
cs.LG2025
Backdoors in DRL: Four Environments Focusing on In-distribution Triggers
Chace Ashcraft, Ted Staley, Josh Carney +4
Backdoor attacks, or trojans, pose a security risk by concealing undesirable behavior in deep neural network models. Open-source neural networks are downloaded from the internet da…
cs.LG2025
Investigating the Treacherous Turn in Deep Reinforcement Learning
Chace Ashcraft, Kiran Karra, Josh Carney +1
The Treacherous Turn refers to the scenario where an artificial intelligence (AI) agent subtly, and perhaps covertly, learns to perform a behavior that benefits itself but is deeme…