1 paper · 1 filter
Hannah Rose Kirk, Iason Gabriel, Chris Summerfield +2
Humans strive to design safe AI systems that align with our goals and remain under our control. However, as AI capabilities advance, we face a new challenge: the emergence of deepe…