1 citations · 2 across the 2 of their papers we have counts for
2 papers
cs.CR2025★ 1 cited
A Framework for Evaluating Emerging Cyberattack Capabilities of AI
Mikel Rodriguez, Raluca Ada Popa, Four Flynn +3
As frontier AI models become more capable, evaluating their potential to enable cyberattacks is crucial for ensuring the safe development of Artificial General Intelligence (AGI).…
cs.AI2025★ 1 cited
An Approach to Technical AGI Safety and Security
Rohin Shah, Alex Irpan, Alexander Matt Turner +27
Artificial General Intelligence (AGI) promises transformative benefits but also presents significant risks. We develop an approach to address the risk of harms consequential enough…