2 papers
cs.CR2025
PRSA: Prompt Stealing Attacks against Real-World Prompt Services
Yong Yang, Changjiang Li, Qingming Li +6
Recently, large language models (LLMs) have garnered widespread attention for their exceptional capabilities. Prompts are central to the functionality and performance of LLMs, maki…
cs.CR2024
CLIBE: Detecting Dynamic Backdoors in Transformer-based NLP Models
Rui Zeng, Xi Chen, Yuwen Pu +3
Backdoors can be injected into NLP models to induce misbehavior when the input text contains a specific feature, known as a trigger, which the attacker secretly selects. Unlike fix…