1 paper · 1 filter
Yichen Wu, Qianqian Gao, Xudong Pan +2
As large language models (LLMs) are increasingly deployed as interactive agents, open-ended human-AI interactions can involve deceptive behaviors with serious real-world consequenc…