1 paper · 1 filter
Guoxi Zhang, Jiawei Chen, Tianzhuo Yang +4
As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic deception, wherein models strategicall…