12 citations · 12 across the 1 of their papers we have counts for
1 paper
Ondrej Bajgar, Jan Horenovsky
If autonomous AI systems are to be reliably safe in novel situations, they will need to incorporate general principles guiding them to recognize and avoid harmful behaviours. Such…