6 citations · 6 across the 1 of their papers we have counts for
1 paper
Yuxia Wang, Haonan Li, Xudong Han +2
With the rapid evolution of large language models (LLMs), new and hard-to-predict harmful capabilities are emerging. This requires developers to be able to identify risks through t…