1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.MA2023★ 1 cited
Welfare Diplomacy: Benchmarking Language Model Cooperation
Gabriel Mukobi, Hannah Erlebach, Niklas Lauffer +3
The growing capabilities and increasingly widespread deployment of AI systems necessitate robust benchmarks for measuring their cooperative capabilities. Unfortunately, most multi-…
cs.CL2023
Towards the Scalable Evaluation of Cooperativeness in Language Models
Alan Chan, Maxime Riché, Jesse Clifton
It is likely that AI systems driven by pre-trained language models (PLMs) will increasingly be used to assist humans in high-stakes interactions with other agents, such as negotiat…