2 papers
cs.CL2022
Distillation-Resistant Watermarking for Model Protection in NLP
Xuandong Zhao, Lei Li, Yu-Xiang Wang
How can we protect the intellectual property of trained NLP models? Modern NLP models are prone to stealing by querying and distilling from their publicly exposed APIs. However, ex…
cs.LG2020
Bullseye Polytope: A Scalable Clean-Label Poisoning Attack with Improved Transferability
Hojjat Aghakhani, Dongyu Meng, Yu-Xiang Wang +2
A recent source of concern for the security of neural networks is the emergence of clean-label dataset poisoning attacks, wherein correctly labeled poison samples are injected into…