5 citations · 15 across the 11 of their papers we have counts for
3 papers · 1 filter
Calibrating LLM Judges: Linear Probes for Fast and Reliable Uncertainty Estimation
Bhaktipriya Radharapu, Eshika Saxena, Kenneth Li +3
As LLM-based judges become integral to industry applications, obtaining well-calibrated uncertainty estimates efficiently has become critical for production deployment. However, ex…
Cultivating Pluralism In Algorithmic Monoculture: The Community Alignment Dataset
Lily Hong Zhang, Smitha Milli, Karen Jusko +12
How can large language models (LLMs) serve users with varying preferences that may conflict across cultural, political, or other dimensions? To advance this challenge, this paper e…
Safety and Fairness for Content Moderation in Generative Models
Susan Hao, Piyush Kumar, Sarah Laszlo +3
With significant advances in generative AI, new technologies are rapidly being deployed with generative components. Generative models are typically trained on large datasets, resul…