Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025
GuardVal: Dynamic Large Language Model Jailbreak Evaluation for Comprehensive Safety Testing
Peiyan Zhang, Haibo Jin, Liying Kang +1
Jailbreak attacks reveal critical vulnerabilities in Large Language Models (LLMs) by causing them to generate harmful or unethical content. Evaluating these threats is particularly…
cs.LG2024
DistDD: Distributed Data Distillation Aggregation through Gradient Matching
Peiran Wang, Haohan Wang
In this paper, we introduce DistDD, a novel approach within the federated learning framework that reduces the need for repetitive communication by distilling data directly on clien…