4 papers
MemCatalyst: Amplifying Data Auditing on Vision-Language Models via Data Poisoning
Xukun Luan, Jinyan Liu, Yuhui Gong +4
Vision-Language models (VLMs) achieve outstanding performance largely due to the amount of training data available on the internet. At the same time, data holders (e.g., artists) u…
Code-Poisoning Property Inference Attacks
Xukun Luan, Yuhui Gong, Gang Zhang +4
The flourishing code hosting platforms and coding agents enable even beginners with private data to build tailored Machine Learning (ML) models using available code quickly. The tr…
Is External Database Protection Static in Retrieval-Augmented Generation? Rethinking Privacy Preservation under Dynamic Queries
Gang Zhang, Mingyu Tian, Xukun Luan +2
Retrieval-augmented generation (RAG) enhances large language models via external document retrieval, but retrieved contexts may leak sensitive information. Current privacy protecti…
VLALeaks: Membership Inference Attacks against Vision-Language-Action Models
Xukun Luan, Jinyan Liu, Xuesong Li +4
Vision-Language-Action (VLA) models enable end-to-end robot control and have garnered widespread attention. However, the memorization of training data inherent to VLA, coupled with…