3 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 3 cited
Improved Techniques for Optimization-Based Jailbreaking on Large Language Models
Xiaojun Jia, Tianyu Pang, Chao Du +5
Large language models (LLMs) are being rapidly developed, and a key component of their widespread deployment is their safety-related alignment. Many red-teaming efforts aim to jail…
cs.LG2024
Contrastive Learning with Negative Sampling Correction
Lu Wang, Chao Du, Pu Zhao +8
As one of the most effective self-supervised representation learning methods, contrastive learning (CL) relies on multiple negative pairs to contrast against each positive pair. In…