3 papers
cs.LG2026
Gecko: Fast Private Inference via Secure Public Encoder Offloading
Cheng'an Wei, Kai Chen, Yue Zhao +2
Private inference protects both user inputs and server models during neural network inference, but existing solutions remain too slow for practical deployment. This motivates recen…
cs.CR2025
LoRAGuard: An Effective Black-box Watermarking Approach for LoRAs
Peizhuo Lv, Yiran Xiahou, Congyi Li +4
LoRA (Low-Rank Adaptation) has achieved remarkable success in the parameter-efficient fine-tuning of large models. The trained LoRA matrix can be integrated with the base model thr…
cs.CL2024
Playing Language Game with LLMs Leads to Jailbreaking
Yu Peng, Zewen Long, Fangming Dong +3
The advent of large language models (LLMs) has spurred the development of numerous jailbreak techniques aimed at circumventing their security defenses against malicious attacks. An…