9 citations · 9 across the 2 of their papers we have counts for
2 papers
cs.CY2022
Aligned with Whom? Direct and social goals for AI systems
Anton Korinek, Avital Balwit
As artificial intelligence (AI) becomes more powerful and widespread, the AI alignment problem - how to ensure that AI systems pursue the goals that we want them to pursue - has ga…
cs.CY2021★ 9 cited
Truthful AI: Developing and governing AI that does not lie
Owain Evans, Owen Cotton-Barratt, Lukas Finnveden +5
In many contexts, lying -- the use of verbal falsehoods to deceive -- is harmful. While lying has traditionally been a human affair, AI systems that make sophisticated verbal state…