activity
20182022
most citedFine-tuning language models to find agreement among humans with diverse preferences

112 citations · 129 across the 5 of their papers we have counts for

collaborators

8 papers

cs.LG2022112 cited

Fine-tuning language models to find agreement among humans with diverse preferences

Michiel A. Bakker, Martin J. Chadwick, Hannah R. Sheahan +8

Recent work in large language modeling (LLMs) has used fine-tuning to align outputs with the preferences of a prototypical user. This work assumes that human preferences are static…

cs.LG20211 cited

Statistical discrimination in learning agents

Edgar A. Duéñez-Guzmán, Kevin R. McKee, Yiran Mao +9

Undesired bias afflicts both human and algorithmic decision making, and may be especially prevalent when information processing trade-offs incentivize the use of heuristics. One pr…

cs.MA2021

Modelling Cooperation in Network Games with Spatio-Temporal Complexity

Michiel A. Bakker, Richard Everett, Laura Weidinger +4

The real world is awash with multi-agent problems that require collective action by self-interested agents, from the routing of packets across a computer network to the management…

cs.LG20191 cited

DADI: Dynamic Discovery of Fair Information with Adversarial Reinforcement Learning

Michiel A. Bakker, Duy Patrick Tu, Humberto Riverón Valdés +4

We introduce a framework for dynamic adversarial discovery of information (DADI), motivated by a scenario where information (a feature set) is used by third parties with unknown ob…

cs.LG201912 cited

Sherlock: A Deep Learning Approach to Semantic Data Type Detection

Madelon Hulsebos, Kevin Hu, Michiel Bakker +5

Correctly detecting the semantic type of data columns is crucial for data science tasks such as automated data cleaning, schema matching, and data discovery. Existing data preparat…

cs.HC20193 cited

VizNet: Towards A Large-Scale Visualization Learning and Benchmarking Repository

Kevin Hu, Neil Gaikwad, Michiel Bakker +7

Researchers currently rely on ad hoc datasets to train automated visualization tools and evaluate the effectiveness of visualization designs. These exemplars often lack the charact…