3 citations · 3 across the 1 of their papers we have counts for
4 papers
The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems
Leon Staufer, Kevin Feng, Kevin Wei +6
Agentic AI systems are increasingly capable of performing professional and personal tasks with limited human involvement. However, tracking these developments is difficult because…
Audit Cards: Contextualizing AI Evaluations
Leon Staufer, Mick Yang, Anka Reuel +1
AI governance frameworks increasingly rely on audits, yet the results of their underlying evaluations require interpretation and context to be meaningfully informative. Even techni…
Mapping Industry Practices to the EU AI Act's GPAI Code of Practice Safety and Security Measures
Lily Stelling, Mick Yang, Rokas Gipiškis +5
This report provides a detailed comparison between the Safety and Security measures proposed in the EU AI Act's General-Purpose AI (GPAI) Code of Practice (Third Draft) and the cur…
Deprecating Benchmarks: Criteria and Framework
Ayrton San Joaquin, Rokas Gipiškis, Leon Staufer +1
As frontier artificial intelligence (AI) models rapidly advance, benchmarks are integral to comparing different models and measuring their progress in different task-specific domai…