13 citations · 17 across the 2 of their papers we have counts for
2 papers
cs.AI2023★ 4 cited
Of Models and Tin Men: A Behavioural Economics Study of Principal-Agent Problems in AI Alignment using Large-Language Models
Steve Phelps, Rebecca Ranson
AI Alignment is often presented as an interaction between a single designer and an artificial agent in which the designer attempts to ensure the agent's behavior is consistent with…
cs.GT2023★ 13 cited
The Machine Psychology of Cooperation: Can GPT models operationalise prompts for altruism, cooperation, competitiveness and selfishness in economic games?
Steve Phelps, Yvan I. Russell
We investigated the capability of the GPT-3.5 large language model (LLM) to operationalize natural language descriptions of cooperative, competitive, altruistic, and self-intereste…