5 papers · 1 filter
The 2026 Singapore Consensus on Global AI Safety Research Priorities
Stephen Casper, Oskar Galeev, Yoshua Bengio +117
Frontier AI capabilities and autonomy are advancing rapidly. A growing number of real-world incidents make a trusted AI ecosystem essential to embracing AI with confidence. The 202…
Eigenism: Ethics for a Human-AI Future
Dan Hendrycks
Our concepts of survival and self-interest were built for single, continuous biological lives. These ideas break down when applied to artificial intelligence, since an AI can be ea…
Virology Capabilities Test (VCT): A Multimodal Virology Q&A Benchmark
Jasper Götting, Pedro Medeiros, Jon G Sanders +6
We present the Virology Capabilities Test (VCT), a large language model (LLM) benchmark that measures the capability to troubleshoot complex virology laboratory protocols. Construc…
Superintelligence Strategy: Expert Version
Dan Hendrycks, Eric Schmidt, Alexandr Wang
Rapid advances in AI are beginning to reshape national security. Destabilizing AI developments could rupture the balance of power and raise the odds of great-power conflict, while…
Beyond Release: Access Considerations for Generative AI Systems
Irene Solaiman, Rishi Bommasani, Dan Hendrycks +4
Generative AI release decisions determine whether system components are made available, but release does not address many other elements that change how users and stakeholders are…