1 citations · 4 across the 13 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2025
From ASR to ASP: Evaluating Prompt Attack Vulnerabilities Against Open-Source LLMs
Jiawen Wang, Pritha Gupta, Ivan Habernal +3
Recent studies demonstrate that Large Language Models (LLMs) are vulnerable to attacks that generate harmful or sensitive outputs. As open-source LLMs are increasingly adopted in h…
cs.CR2023
DP-BART for Privatized Text Rewriting under Local Differential Privacy
Timour Igamberdiev, Ivan Habernal
Privatized text rewriting with local differential privacy (LDP) is a recent approach that enables sharing of sensitive textual documents while formally guaranteeing privacy protect…