From the 1 of 10 linked papers with an AI index.
1 citations · 1 across the 6 of their papers we have counts for
10 papers
DiffSafeMerge: Mitigating Backdoor Inheritance in Diffusion Model Merging
Jiayang Zhang, Ji Guo, Jiachen Li +2
Unconditional diffusion checkpoint merging assumes benign sources, yet a compromised public checkpoint can transfer a dormant backdoor while clean generation appears normal. Mitiga…
EvoTrustRAG: Evolution-Aware Conflict Attribution and Evidence Handling for Reliable Retrieval-Augmented Generation
Xi Nie, Hongwei Li, Shenghao Wu +3
Retrieval-Augmented Generation (RAG) improves the factuality of large language models with external knowledge, yet conflicting evidence remains a fundamental challenge in dynamic a…
InkShield: Writing Style Protection Against Unauthorized Handwriting Mimicry
Jian Xiong, Wenbo Jiang, Zihan Wang +4
InkShield introduces a proactive defense that adds subtle, stroke‑confined perturbations to handwritten reference images, making it harder for handwriting generators to mimic a wri…
UTOPIA: Unlearnable Tabular Data via Decoupled Shortcut Embedding
Jiaming He, Fuming Luo, Hongwei Li +5
Unlearnable examples (UE) have emerged as a practical mechanism to prevent unauthorized model training on private vision data, while extending this protection to tabular data is no…
ConfGuard: A Simple and Effective Backdoor Detection for Large Language Models
Zihan Wang, Rui Zhang, Hongwei Li +4
Backdoor attacks pose a significant threat to Large Language Models (LLMs), where adversaries can embed hidden triggers to manipulate LLM's outputs. Most existing defense methods,…
MPMA: Preference Manipulation Attack Against Model Context Protocol
Zihan Wang, Rui Zhang, Yu Liu +5
Model Context Protocol (MCP) standardizes interface mapping for large language models (LLMs) to access external data and tools, which revolutionizes the paradigm of tool selection…