From the 1 of 8 linked papers with an AI index.
8 papers
AMT-X: Phase-Structured Multi-Turn Red-Teaming with Checklist-Gated Evaluation
Yi Ting Shen, Kentaroh Toyoda, Alex Leung
The paper introduces AMT‑X, a framework that conducts multi‑turn red‑team attacks on large language models using a phase‑structured state machine and evaluates success with a multi…
The 2026 Singapore Consensus on Global AI Safety Research Priorities
Stephen Casper, Oskar Galeev, Yoshua Bengio +117
Frontier AI capabilities and autonomy are advancing rapidly. A growing number of real-world incidents make a trusted AI ecosystem essential to embracing AI with confidence. The 202…
The Insurability Frontier of AI Risk: Mapping Threats to Affirmative Coverage, Silent Exposures, and Exclusions
Alex Leung, Rex Zhang, Ervin Ling +2
The rapid diffusion of agentic AI has created a new coverage problem for commercial insurance: some AI-mediated losses are now affirmatively insured, some create silent-AI exposure…
From Control Boundary to Insurance Claim: Reconstructing AI-Mediated Losses Through the CER Framework
Alex Leung, Rex Zhang, Kentaroh Toyoda +1
AI losses that arise through an insured organization's generative or agentic AI system require state reconstruction, not merely event reconstruction, because the relevant state cha…
IPI-proxy: An Intercepting Proxy for Red-Teaming Web-Browsing AI Agents Against Indirect Prompt Injection
Chia-Pei, Chen, Kentaroh Toyoda +2
Web-browsing AI agents are increasingly deployed in enterprise settings under strict whitelists of approved domains, yet adversaries can still influence them by embedding hidden in…
AI Identity: Standards, Gaps, and Research Directions for AI Agents
Takumi Otsuka, Kentaroh Toyoda, Alex Leung
AI agents are now running real transactions, workflows, and sub-agent chains across organizational boundaries without continuous human supervision. This creates a problem no curren…