155 citations · 170 across the 12 of their papers we have counts for
1 paper · 1 filter
Alessio Benavoli, Alessandro Facchini, Marco Zaffalon
How can we ensure that AI systems are aligned with human values and remain safe? We can study this problem through the frameworks of the AI assistance and the AI shutdown games. Th…