1 paper · 1 filter
Anthony Costarelli, Mat Allen, Severin Field
As Large Language Models (LLMs) become increasingly integrated into our daily lives, the potential harms from deceptive behavior underlie the need for faithfully interpreting their…