4 papers
Backchaining Loss of Control Mitigations from Mission-Specific Benchmarks in National Security
Matteo Pistillo, Samantha Faraone, Joshua Herman
Affordances and permissions are promising and timely safety levers for mitigating Loss of Control (LoC) threats in high-stakes deployment contexts, such as national security. Deplo…
Internal Deployment in the AI Act
Matteo Pistillo
This memorandum analyzes and stress-tests arguments in favor and against the inclusion of internal deployment within the scope of the European Union Artificial Intelligence Act (AI…
Towards Frontier Safety Policies Plus
Matteo Pistillo
This paper examines the state of affairs on Frontier Safety Policies in light of capability progress and growing expectations held by government actors and AI safety researchers fr…
Defending Compute Thresholds Against Legal Loopholes
Matteo Pistillo, Pablo Villalobos
Existing legal frameworks on AI rely on training compute thresholds as a proxy to identify potentially-dangerous AI models and trigger increased regulatory attention. In the United…