3 papers
cs.CL2026
Selective Knowledge Edit Reversal via Gated Singular Vector Shrinkage
Weifeng Jiang, Ruirui Chen, Qianren Mao +3
Knowledge editing provides an efficient way to update factual knowledge in large language models. However, malicious edits may introduce safety risks, making it necessary to revers…
cs.AI2024
AI Safety Landscape for Large Language Models: Taxonomy, State-of-the-art, and Future Directions
Chen Chen, Xueluan Gong, Ziyao Liu +3
AI Safety is an emerging area of critical importance to the safe adoption and deployment of AI systems. With the rapid proliferation of AI and especially with the recent advancemen…
cs.CR2024
Guaranteeing Data Privacy in Federated Unlearning with Dynamic User Participation
Ziyao Liu, Yu Jiang, Weifeng Jiang +3
Federated Unlearning (FU) is gaining prominence for its capability to eliminate influences of Federated Learning (FL) users' data from trained global FL models. A straightforward F…