1 paper
Akshat Naik, Emma Gouné, Patrick Quinn +4
As Large Language Model (LLM) agents become more widespread, associated misalignment risks increase. While prior research has studied agents' ability to produce harmful outputs or…