4 papers
NADD: Amplifying Noise for Effective Diffusion-based Adversarial Purification
David D. Nguyen, The-Anh Ta, Yansong Gao +1
The strategy of combining diffusion-based generative models with classifiers continues to demonstrate state-of-the-art performance on adversarial robustness benchmarks. Known as ad…
Alert-ME: An Explainability-Driven Defense Against Adversarial Examples in Transformer-Based Text Classification
Bushra Sabir, Yansong Gao, Alsharif Abuadbba +1
Transformer-based text classifiers such as BERT, RoBERTa, T5, and GPT have shown strong performance in natural language processing tasks but remain vulnerable to adversarial exampl…
Large Language Model Adversarial Landscape Through the Lens of Attack Objectives
Nan Wang, Kane Walter, Yansong Gao +1
Large Language Models (LLMs) represent a transformative leap in artificial intelligence, enabling the comprehension, generation, and nuanced interaction with human language on an u…
Comprehensive Evaluation of Cloaking Backdoor Attacks on Object Detector in Real-World
Hua Ma, Alsharif Abuadbba, Yansong Gao +2
The exploration of backdoor vulnerabilities in object detectors, particularly in real-world scenarios, remains limited. A significant challenge lies in the absence of a natural phy…