3 papers
cs.CL2024
SPICA: Retrieving Scenarios for Pluralistic In-Context Alignment
Quan Ze Chen, K. J. Kevin Feng, Chan Young Park +1
When different groups' values differ, one approach to model alignment is to steer models at inference time towards each group's preferences. However, techniques like in-context lea…
cs.HC2024
Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment
Andrew Konya, Aviv Ovadya, Kevin Feng +4
We introduce a method to measure the alignment between public will and language model (LM) behavior that can be applied to fine-tuning, online oversight, and pre-release safety che…
cs.CY2024
Democratic AI is Possible. The Democracy Levels Framework Shows How It Might Work
Aviv Ovadya, Kyle Redman, Luke Thorburn +11
This position paper argues that effectively "democratizing AI" requires democratic governance and alignment of AI, and that this is particularly valuable for decisions with systemi…