1 citations · 1 across the 3 of their papers we have counts for
1 paper · 1 filter
Vedant Bhasin, Matthew Yudin, Razvan Stefanescu +1
Trojan backdoors can be injected into large language models at various stages, including pretraining, fine-tuning, and in-context learning, posing a significant threat to the model…