1 paper
Atharvan Dogra, Krishna Pillutla, Ameet Deshpande +5
We explore the ability of large language models (LLMs) to engage in subtle deception through strategically phrasing and intentionally manipulating information. This harmful behavio…