1 paper · 2 filters
Nay Myat Min, Long H. Pham, Yige Li +1
Large Language Models (LLMs) are vulnerable to backdoor attacks that manipulate outputs via hidden triggers. Existing defense methods--designed for vision/text classification tasks…