55 citations · 77 across the 3 of their papers we have counts for
Showing cs.AIShow all
2 papers · 1 filter
cs.AI2023
AI Systems of Concern
Kayla Matteucci, Shahar Avin, Fazl Barez +1
Concerns around future dangers from advanced AI often centre on systems hypothesised to have intrinsic characteristics such as agent-like behaviour, strategic awareness, and long-r…
cs.AI2021★ 55 cited
Filling gaps in trustworthy development of AI
Shahar Avin, Haydn Belfield, Miles Brundage +9
The range of application of artificial intelligence (AI) is vast, as is the potential for harm. Growing awareness of potential risks from AI systems has spurred action to address t…