6 citations · 11 across the 4 of their papers we have counts for
4 papers
More than Marketing? On the Information Value of AI Benchmarks for Practitioners
Amelia Hardy, Anka Reuel, Kiana Jafari Meimandi +6
Public AI benchmark results are widely broadcast by model developers as indicators of model quality within a growing and competitive market. However, these advertised scores do not…
GPAI Evaluations Standards Taskforce: Towards Effective AI Governance
Patricia Paskov, Lukas Berglund, Everett Smith +1
General-purpose AI evaluations have been proposed as a promising way of identifying and mitigating systemic risks posed by AI development and deployment. While GPAI evaluations pla…
Position Paper: Technical Research and Talent is Needed for Effective AI Governance
Anka Reuel, Lisa Soder, Ben Bucknall +1
In light of recent advancements in AI capabilities and the increasingly widespread integration of AI systems into society, governments worldwide are actively seeking to mitigate th…
Towards Publicly Accountable Frontier LLMs: Building an External Scrutiny Ecosystem under the ASPIRE Framework
Markus Anderljung, Everett Thornton Smith, Joe O'Brien +7
With the increasing integration of frontier large language models (LLMs) into society and the economy, decisions related to their training, deployment, and use have far-reaching im…