1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Roman Belaire, Arunesh Sinha, Pradeep Varakantham
Red teaming is critical for identifying vulnerabilities and building trust in current LLMs. However, current automated methods for Large Language Models (LLMs) rely on brittle prom…