4 citations · 4 across the 2 of their papers we have counts for
2 papers
cs.LG2025
Full-Stack Alignment: Co-Aligning AI and Institutions with Thick Models of Value
Joe Edelman, Tan Zhi-Xuan, Ryan Lowe +30
Beneficial societal outcomes cannot be guaranteed by aligning individual AI systems with the intentions of their operators or users. Even an AI system that is perfectly aligned to…
cs.CY2025★ 4 cited
Characterizing AI Agents for Alignment and Governance
Atoosa Kasirzadeh, Iason Gabriel
The creation of effective governance mechanisms for AI agents requires a deeper understanding of their core properties and how these properties relate to questions surrounding the…