2 papers
cs.CY2026
The 2026 Singapore Consensus on Global AI Safety Research Priorities
Stephen Casper, Oskar Galeev, Yoshua Bengio +117
Frontier AI capabilities and autonomy are advancing rapidly. A growing number of real-world incidents make a trusted AI ecosystem essential to embracing AI with confidence. The 202…
cs.CL2025
CoPE: A Small Language Model for Steerable and Scalable Content Labeling
Samidh Chakrabarti, David Willner, Kevin Klyman +3
This paper details the methodology behind CoPE, a policy-steerable small language model capable of fast and accurate content labeling. We present a novel training curricula called…