1 paper · 1 filter
Julian Rodemann, Esteban Garces Arias, Christoph Luther +2
Empirical human-AI alignment aims to make AI systems act in line with observed human behavior. While noble in its goals, we argue that empirical alignment can inadvertently introdu…