4 papers
CleanGen: Mitigating Backdoor Attacks for Generation Tasks in Large Language Models
Yuetai Li, Zhangchen Xu, Fengqing Jiang +4
The remarkable performance of large language models (LLMs) in generation tasks has enabled practitioners to leverage publicly available models to power custom applications, such as…
Who is Responsible? Explaining Safety Violations in Multi-Agent Cyber-Physical Systems
Luyao Niu, Hongchao Zhang, Dinuka Sahabandu +3
Multi-agent cyber-physical systems are present in a variety of applications. Agent decision-making can be affected due to errors induced by uncertain, dynamic operating environment…
A Method for Fast Autonomy Transfer in Reinforcement Learning
Dinuka Sahabandu, Bhaskar Ramasubramanian, Michail Alexiou +3
This paper introduces a novel reinforcement learning (RL) strategy designed to facilitate rapid autonomy transfer by utilizing pre-trained critic value functions from multiple envi…
Game of Trojans: Adaptive Adversaries Against Output-based Trojaned-Model Detectors
Dinuka Sahabandu, Xiaojun Xu, Arezoo Rajabi +4
We propose and analyze an adaptive adversary that can retrain a Trojaned DNN and is also aware of SOTA output-based Trojaned model detectors. We show that such an adversary can ens…