2 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CR2025
HoneypotNet: Backdoor Attacks Against Model Extraction
Yixu Wang, Tianle Gu, Yan Teng +2
Model extraction attacks are one type of inference-time attacks that approximate the functionality and performance of a black-box victim model by launching a certain number of quer…
cs.CL2024★ 2 cited
MEOW: MEMOry Supervised LLM Unlearning Via Inverted Facts
Tianle Gu, Kexin Huang, Ruilin Luo +4
Large Language Models (LLMs) can memorize sensitive information, raising concerns about potential misuse. LLM Unlearning, a post-hoc approach to remove this information from traine…