2 papers
cs.LG2026
Adaptive Pluralistic Alignment: A pipeline for dynamic artificial democracy
Rachel Freedman
Prevailing alignment methods target a fixed set of preferences and therefore risk forcing value lock-in as societal norms evolve over time. We introduce Adaptive Pluralistic Alignm…
cs.AI2026
Active teacher selection for reward learning
Rachel Freedman, Justin Svegliato, Kyle Wray +1
Reward learning techniques enable machine learning systems to learn objectives from human feedback. A core limitation of these systems is their assumption that all feedback comes f…