17 citations · 20 across the 23 of their papers we have counts for
Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Functional Subspace Watermarking for Large Language Models
Zikang Ding, Junhao Li, Suling Wu +3
Model watermarking utilizes internal representations to protect the ownership of large language models (LLMs). However, these features inevitably undergo complex distortions during…
cs.CR2025
Backdooring CLIP through Concept Confusion
Lijie Hu, Junchi Liao, Weimin Lyu +5
Backdoor attacks pose a serious threat to deep learning models by allowing adversaries to implant hidden behaviors that remain dormant on clean inputs but are maliciously triggered…