2 citations · 9 across the 18 of their papers we have counts for
Showing 2024Show all
3 papers · 1 filter
cs.CR2024★ 2 cited
Eguard: Defending LLM Embeddings Against Inversion Attacks via Text Mutual Information Optimization
Tiantian Liu, Hongwei Yao, Feng Lin +3
Embeddings have become a cornerstone in the functionality of large language models (LLMs) due to their ability to transform text data into rich, dense numerical representations tha…
cs.CR2024★ 1 cited
ShadowCode: Towards (Automatic) External Prompt Injection Attack against Code LLMs
Yuchen Yang, Yiming Li, Hongwei Yao +5
Recent advancements have led to the widespread adoption of code-oriented large language models (Code LLMs) for programming tasks. Despite their success in deployment, their securit…
cs.CR2024★ 1 cited
Explanation as a Watermark: Towards Harmless and Multi-bit Model Ownership Verification via Watermarking Feature Attribution
Shuo Shao, Yiming Li, Hongwei Yao +3
Ownership verification is currently the most critical and widely adopted post-hoc method to safeguard model copyright. In general, model owners exploit it to identify whether a giv…