activity
20122025
most citedBadChain: Backdoor Chain-of-Thought Prompting for Large Language Models

8 citations · 25 across the 17 of their papers we have counts for

collaborators

17 papers

cs.AI2025

SafeChain: Safety of Language Models with Long Chain-of-Thought Reasoning Capabilities

Fengqing Jiang, Zhangchen Xu, Yuetai Li +5

Emerging large reasoning models (LRMs), such as DeepSeek-R1 models, leverage long chain-of-thought (CoT) reasoning to generate structured intermediate steps, enhancing their reason…

eess.SY2024

Modeling and Designing Non-Pharmaceutical Interventions in Epidemics: A Submodular Approach

Shiyu Cheng, Luyao Niu, Bhaskar Ramasubramanian +2

This paper considers the problem of designing non-pharmaceutical intervention (NPI) strategies, such as masking and social distancing, to slow the spread of a viral epidemic. We fo…

cs.LG2024

A Method for Fast Autonomy Transfer in Reinforcement Learning

Dinuka Sahabandu, Bhaskar Ramasubramanian, Michail Alexiou +3

This paper introduces a novel reinforcement learning (RL) strategy designed to facilitate rapid autonomy transfer by utilizing pre-trained critic value functions from multiple envi…

cs.CR2024

ACE: A Model Poisoning Attack on Contribution Evaluation Methods in Federated Learning

Zhangchen Xu, Fengqing Jiang, Luyao Niu +3

In Federated Learning (FL), a set of clients collaboratively train a machine learning model (called global model) without sharing their local training data. The local training data…

cs.RO2024

Fault Tolerant Neural Control Barrier Functions for Robotic Systems under Sensor Faults and Attacks

Hongchao Zhang, Luyao Niu, Andrew Clark +1

Safety is a fundamental requirement of many robotic systems. Control barrier function (CBF)-based approaches have been proposed to guarantee the safety of robotic systems. However,…

cs.CR2024

Game of Trojans: Adaptive Adversaries Against Output-based Trojaned-Model Detectors

Dinuka Sahabandu, Xiaojun Xu, Arezoo Rajabi +4

We propose and analyze an adaptive adversary that can retrain a Trojaned DNN and is also aware of SOTA output-based Trojaned model detectors. We show that such an adversary can ens…