1 paper · 1 filter
Saurabh Ranjan, Konstantina Sokratous, Brian Odegaard
A conversational AI that cannot tell its own output from what a user said will treat its own mistakes as user-provided facts. In humans, this capacity is called reality monitoring,…