2 papers
cs.AI2026
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
Sheza Munir, Benjamin Mah, Krisha Kalsi +5
In machine learning, "ground truth" refers to the assumed correct labels used to train and evaluate models. However, the foundational "ground truth" paradigm rests on a positivisti…
cs.AI2025
Reflective Verbal Reward Design for Pluralistic Alignment
Carter Blair, Kate Larson, Edith Law
AI agents are commonly aligned with "human values" through reinforcement learning from human feedback (RLHF), where a single reward model is learned from aggregated human feedback…