2 papers
cs.CR2025
Model-agnostic clean-label backdoor mitigation in cybersecurity environments
Giorgio Severi, Simona Boboila, John Holodnak +4
The training phase of machine learning models is a delicate step, especially in cybersecurity contexts. Recent research has surfaced a series of insidious training-time attacks tha…
cs.CL2025
Trojan Detection Through Pattern Recognition for Large Language Models
Vedant Bhasin, Matthew Yudin, Razvan Stefanescu +1
Trojan backdoors can be injected into large language models at various stages, including pretraining, fine-tuning, and in-context learning, posing a significant threat to the model…