1 paper · 1 filter
Christopher Kao, Vanshika Vats, James Davis
Large Language Model (LLM) agents are increasingly used in many applications, raising concerns about their safety. While previous work has shown that LLMs can deceive in controlled…