Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Cordyceps: Covert Control Attacks on LLMs via Data Poisoning
Zedian Shao, Charles Fleming, Teodora Baluta
Large language models (LLMs) are often fine-tuned on uncurated text datasets that adversaries can poison. Existing poisoning attacks primarily rely on fixed trigger phrases that de…
cs.CR2025
Watermark Robustness and Radioactivity May Be at Odds in Federated Learning
Leixu Huang, Zedian Shao, Teodora Baluta
Federated learning (FL) enables fine-tuning large language models (LLMs) across distributed data sources. As these sources increasingly include LLM-generated text, provenance track…