2 papers
cs.CL2025
ChineseSafe: A Chinese Benchmark for Evaluating Safety in Large Language Models
Hengxiang Zhang, Hongfu Gao, Qiang Hu +7
With the rapid development of Large language models (LLMs), understanding the capabilities of LLMs in identifying unsafe content has become increasingly important. While previous w…
cs.LG2024
Defending Membership Inference Attacks via Privacy-aware Sparsity Tuning
Qiang Hu, Hengxiang Zhang, Hongxin Wei
Over-parameterized models are typically vulnerable to membership inference attacks, which aim to determine whether a specific sample is included in the training of a given model. P…