2 papers
cs.CR2026
AI Security Leaderboard: Methodology, Results and Minimal Standard
Jasper Timm, Lukas Struppek, Ziwei Xu +12
The AI Security Leaderboard is an independent benchmark that ranks the safeguards of frontier AI models from least to most secure. It tests models against the FARAI Minimal Stan…
cs.SI2026
When Agents Talk: Discourse, Manipulation, and Risk in an Agentic Social Network
10a Labs, :, Grace Cheong +22
AI agents are increasingly interacting within shared online environments, creating new operational security risks. We analyze activity on Moltbook, a Reddit-style social platform w…