5 papers · 1 filter
Botender: Supporting Communities in Collaboratively Designing AI Agents through Case-Based Provocations
Tzu-Sheng Kuo, Sophia Liu, Quan Ze Chen +4
AI agents, or bots, serve important roles in online communities. However, they are often designed by outsiders or a few tech-savvy members, leading to bots that may not align with…
Chain of Alignment: Integrating Public Will with Expert Intelligence for Language Model Alignment
Andrew Konya, Aviv Ovadya, Kevin Feng +4
We introduce a method to measure the alignment between public will and language model (LM) behavior that can be applied to fine-tuning, online oversight, and pre-release safety che…
PolicyCraft: Supporting Collaborative and Participatory Policy Design through Case-Grounded Deliberation
Tzu-Sheng Kuo, Quan Ze Chen, Amy X. Zhang +3
Community and organizational policies are typically designed in a top-down, centralized fashion, with limited input from impacted stakeholders. This can result in policies that are…
Policy Prototyping for LLMs: Pluralistic Alignment via Interactive and Collaborative Policymaking
K. J. Kevin Feng, Inyoung Cheong, Quan Ze Chen +1
Emerging efforts in AI alignment seek to broaden participation in shaping model behavior by eliciting and integrating collective input into a policy for model finetuning. While plu…
End User Authoring of Personalized Content Classifiers: Comparing Example Labeling, Rule Writing, and LLM Prompting
Leijie Wang, Kathryn Yurechko, Pranati Dani +2
Existing tools for laypeople to create personal classifiers often assume a motivated user working uninterrupted in a single, lengthy session. However, users tend to engage with soc…