activity
20222024
most citedBadChain: Backdoor Chain-of-Thought Prompting for Large Language Models

8 citations · 9 across the 6 of their papers we have counts for

collaborators

6 papers

cs.CR2024

Game of Trojans: Adaptive Adversaries Against Output-based Trojaned-Model Detectors

Dinuka Sahabandu, Xiaojun Xu, Arezoo Rajabi +4

We propose and analyze an adaptive adversary that can retrain a Trojaned DNN and is also aware of SOTA output-based Trojaned model detectors. We show that such an adversary can ens…

cs.LG2024

Double-Dip: Thwarting Label-Only Membership Inference Attacks with Transfer Learning and Randomization

Arezoo Rajabi, Reeya Pimple, Aiswarya Janardhanan +3

Transfer learning (TL) has been demonstrated to improve DNN model performance when faced with a scarcity of training samples. However, the suitability of TL as a solution to reduce…

cs.CR20248 cited

BadChain: Backdoor Chain-of-Thought Prompting for Large Language Models

Zhen Xiang, Fengqing Jiang, Zidi Xiong +3

Large language models (LLMs) are shown to benefit from chain-of-thought (COT) prompting, particularly when tackling tasks that require systematic reasoning processes. On the other…

cs.CR20231 cited

MDTD: A Multi Domain Trojan Detector for Deep Neural Networks

Arezoo Rajabi, Surudhi Asokraj, Fengqing Jiang +4

Machine learning models that use deep neural networks (DNNs) are vulnerable to backdoor attacks. An adversary carrying out a backdoor attack embeds a predefined perturbation called…

cs.AI2023

Risk-Aware Distributed Multi-Agent Reinforcement Learning

Abdullah Al Maruf, Luyao Niu, Bhaskar Ramasubramanian +2

Autonomous cyber and cyber-physical systems need to perform decision-making, learning, and control in unknown environments. Such decision-making can be sensitive to multiple factor…

cs.LG2022

Game of Trojans: A Submodular Byzantine Approach

Dinuka Sahabandu, Arezoo Rajabi, Luyao Niu +3

Machine learning models in the wild have been shown to be vulnerable to Trojan attacks during training. Although many detection mechanisms have been proposed, strong adaptive attac…