5 papers
Silent Sabotage: Internal State Triggered Backdoor Attacks on LLM-Powered Robotic Systems
Doniyorkhon Obidov, Shivayogi Akki, Tan Chen +1
The integration of Large Language Models (LLMs) into robotic control systems is enabling a new generation of autonomous agents capable of complex reasoning and planning. While this…
Dynamic Deep Prompt Optimization for Defending Against Jailbreak Attacks on LLMs
Doniyorkhon Obidov, Honggang Yu, Xiaolong Guo +1
Large Language Models (LLMs) demonstrate impressive capabilities across many applications but remain vulnerable to jailbreak attacks, which elicit harmful or unintended content. Wh…
StepTrigger: Contact-State-Triggered Backdoor Attacks on VLM-Powered Legged Robots
Jiageng Zhang, Doniyorkhon Obidov, Kaichen Yang
Large language models and vision-language models are increasingly used as high-level planners in robotic systems, using task goals and sensor summaries to select navigation or mani…
LoRAScan: Detecting Backdoor Prompts in Low-Rank Adapters for Large Language Models via Down-Projection Activation Spikes
Doniyorkhon Obidov, Honggang Yu, Xiaolong Guo +1
Low-rank adaptation (LoRA) enables efficient specialization and distribution of large language models through compact adapters. However, untrusted adapters introduce a supply-chain…
Genotypic Triggers: Exposing Pharmacogenomic Blind Spots via Host-Specific Backdoors in Generative Antimicrobial Peptide Models
Doniyorkhon Obidov, Xiaolong Guo, Yonghui Li +1
Large Language Models (LLMs) have accelerated drug discovery, particularly in the automated design of antimicrobial peptides (AMPs). However, current validation pipelines for pepti…