1 paper · 1 filter
Christos Ziakas, Nicholas Loo, Nishita Jain +1
Automated red-teaming has emerged as a scalable approach for auditing Large Language Models (LLMs) prior to deployment, yet existing approaches lack mechanisms to efficiently adapt…