1 paper
Teun van der Weij, Felix Hofstätter, Ollie Jaffe +2
Trustworthy capability evaluations are crucial for ensuring the safety of AI systems, and are becoming a key component of AI regulation. However, the developers of an AI system, or…