3 papers
cs.CR2025
ExpShield: Safeguarding Web Text from Unauthorized Crawling and LLM Exploitation
Ruixuan Liu, Toan Tran, Tianhao Wang +3
As large language models increasingly memorize web-scraped training content, they risk exposing copyrighted or private information. Existing protections require compliance from cra…
cs.AI2024
AI-Compass: A Comprehensive and Effective Multi-module Testing Tool for AI Systems
Zhiyu Zhu, Zhibo Jin, Hongsheng Hu +5
AI systems, in particular with deep learning techniques, have demonstrated superior performance for various real-world applications. Given the need for tailored optimization in spe…
cs.CR2024
Releasing Malevolence from Benevolence: The Menace of Benign Data on Machine Unlearning
Binhao Ma, Tianhang Zheng, Hongsheng Hu +5
Machine learning models trained on vast amounts of real or synthetic data often achieve outstanding predictive performance across various domains. However, this utility comes with…