4 papers
SRAF: Stealthy and Robust Adversarial Fingerprint for Copyright Verification of Large Language Models
Zhebo Wang, Zhenhua Xu, Maike Li +4
The protection of Intellectual Property (IP) for Large Language Models (LLMs) has become a critical concern as model theft and unauthorized commercialization escalate. While advers…
Latent Fusion Jailbreak: Blending Harmful and Harmless Representations to Elicit Unsafe LLM Outputs
Wenpeng Xing, Mohan Li, Bohan Yang +5
Safety-aligned large language models can still be manipulated through white-box interventions that modify their internal representations. We introduce Latent Fusion Jailbreak (LFJ)…
Local Distance Query with Differential Privacy
Weihong Sheng, Jiajun Chen, Bin Cai +3
Differential Privacy (DP) is commonly employed to safeguard graph analysis or publishing. Distance, a critical factor in graph analysis, is typically handled using curator DP, wher…
Differentially Private Distance Query with Asymmetric Noise
Weihong Sheng, Jiajun Chen, Chunqiang Hu +3
With the growth of online social services, social information graphs are becoming increasingly complex. Privacy issues related to analyzing or publishing on social graphs are also…