11 citations · 14 across the 3 of their papers we have counts for
3 papers
cs.CY2024★ 1 cited
The Reasonable Person Standard for AI
Sunayana Rane
As AI systems are increasingly incorporated into domains where human behavior has set the norm, a challenge for AI governance and AI alignment research is to regulate their behavio…
cs.LG2024★ 11 cited
Concept Alignment
Sunayana Rane, Polyphony J. Bruna, Ilia Sucholutsky +2
Discussion of AI alignment (alignment between humans and AI systems) has focused on value alignment, broadly referring to creating AI systems that share human values. We argue that…
cs.AI2023★ 2 cited
Concept Alignment as a Prerequisite for Value Alignment
Sunayana Rane, Mark Ho, Ilia Sucholutsky +1
Value alignment is essential for building AI systems that can safely and reliably interact with people. However, what a person values -- and is even capable of valuing -- depends o…