Showing cs.CRShow all
3 papers · 1 filter
cs.CR2025
TokenMark: A Modality-Agnostic Watermark for Pre-trained Transformers
Hengyuan Xu, Liyao Xiang, Borui Yang +3
Watermarking is a critical tool for model ownership verification. However, existing watermarking techniques are often designed for specific data modalities and downstream tasks, wi…
cs.CR2025
HoneypotNet: Backdoor Attacks Against Model Extraction
Yixu Wang, Tianle Gu, Yan Teng +2
Model extraction attacks are one type of inference-time attacks that approximate the functionality and performance of a black-box victim model by launching a certain number of quer…
cs.CR2024
Special Characters Attack: Toward Scalable Training Data Extraction From Large Language Models
Yang Bai, Ge Pei, Jindong Gu +2
Large language models (LLMs) have achieved remarkable performance on a wide range of tasks. However, recent studies have shown that LLMs can memorize training data and simple repea…