2 papers
cs.CY2026
AI, Digital Platforms, and the New Systemic Risk
Philipp Hacker, Lilian Edwards, Atoosa Kasirzadeh
As artificial intelligence (AI) becomes increasingly embedded in digital, social, and institutional infrastructures, and AI and platforms are merged into hybrid structures, systemi…
cs.LG2024
Foundational Challenges in Assuring Alignment and Safety of Large Language Models
Usman Anwar, Abulhair Saparov, Javier Rando +39
This work identifies 18 foundational challenges in assuring the alignment and safety of large language models (LLMs). These challenges are organized into three different categories…