1 paper
Zeyang Yue, Chenfei Yan, Feifei Zhao +5
Whether Large Language Models (LLMs) exhibit covert psychological manipulation in complex human-AI interactions has garnered increasing safety concerns. However, existing AI safety…