112 citations · 138 across the 4 of their papers we have counts for
4 papers
Fine-tuning language models to find agreement among humans with diverse preferences
Michiel A. Bakker, Martin J. Chadwick, Hannah R. Sheahan +8
Recent work in large language modeling (LLMs) has used fine-tuning to align outputs with the preferences of a prototypical user. This work assumes that human preferences are static…
The Good Shepherd: An Oracle Agent for Mechanism Design
Jan Balaguer, Raphael Koster, Christopher Summerfield +1
From social networks to traffic routing, artificial learning agents are playing a central role in modern institutions. We must therefore understand how to leverage these systems to…
HCMD-zero: Learning Value Aligned Mechanisms from Data
Jan Balaguer, Raphael Koster, Ari Weinstein +4
Artificial learning agents are mediating a larger and larger number of interactions among humans, firms, and organizations, and the intersection between mechanism design and machin…
PonderNet: Learning to Ponder
Andrea Banino, Jan Balaguer, Charles Blundell
In standard neural networks the amount of computation used grows with the size of the inputs, but not with the complexity of the problem being learnt. To overcome this limitation w…