3 citations · 3 across the 6 of their papers we have counts for
1 paper · 2 filters
Yu Yan, Sheng Sun, Mingfeng Li +6
To prevent the misuse of Large Language Models (LLMs) for malicious purposes, numerous efforts have been made to develop the safety alignment mechanisms of LLMs. However, as multip…