2 papers
cs.CR2025
A Framework for Evaluating Emerging Cyberattack Capabilities of AI
Mikel Rodriguez, Raluca Ada Popa, Four Flynn +3
As frontier AI models become more capable, evaluating their potential to enable cyberattacks is crucial for ensuring the safe development of Artificial General Intelligence (AGI).…
cs.AI2025
An Approach to Technical AGI Safety and Security
Rohin Shah, Alex Irpan, Alexander Matt Turner +27
Artificial General Intelligence (AGI) promises transformative benefits but also presents significant risks. We develop an approach to address the risk of harms consequential enough…