72 citations · 78 across the 10 of their papers we have counts for
1 paper · 2 filters
Wesley Hanwen Deng, Sunnie S. Y. Kim, Akshita Jha +4
Recent developments in AI governance and safety research have called for red-teaming methods that can effectively surface potential risks posed by AI models. Many of these calls ha…