3 citations · 3 across the 2 of their papers we have counts for
1 paper · 1 filter
Dorian Quelle, Lisa-Maria Neudert, Jonathan Bright +1
In this paper we present an active, constantly updated AI benchmark which measures the integrity of frontier language models against being co-opted for use by authoritarian state "…