1 paper · 1 filter
Tim Rogers, Ben Teehankee
This paper examines a critical yet unexplored dimension of the AI alignment problem: the potential for Large Language Models (LLMs) to inherit and amplify existing misalignments be…