Showing cs.CYShow all
2 papers · 1 filter
cs.CY2026
When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games
Jerick Shi, Terry Jingcheng Zhang, Bernhard Schölkopf +2
As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents that publicly commit to actions w…
cs.CY2026
Cheap Talk, Empty Promise: Frontier LLMs easily break public promises for self-interest
Jerick Shi, Terry Jingcheng Zhang, Zhijing Jin +1
Large language models are increasingly deployed as autonomous agents in multi-agent settings where they communicate intentions and take consequential actions with limited human ove…