4 citations · 5 across the 3 of their papers we have counts for
3 papers · 1 filter
Risk Reporting for Developers' Internal AI Model Use
Oscar Delaney, Sambhav Maheshwari, Joe O'Brien +2
Frontier AI companies first deploy their most advanced models internally, for weeks or months of safety testing, evaluation, and iteration, before a possible public release. For ex…
Mapping Technical Safety Research at AI Companies: A literature review and incentives analysis
Oscar Delaney, Oliver Guest, Zoe Williams
As AI systems become more advanced, concerns about large-scale risks from misuse or accidents have grown. This report analyzes the technical research into safe AI development being…
Coordinated Disclosure of Dual-Use Capabilities: An Early Warning System for Advanced AI
Joe O'Brien, Shaun Ee, Jam Kraprayoon +3
Advanced AI systems may be developed which exhibit capabilities that present significant risks to public safety or security. They may also exhibit capabilities that may be applied…