4 papers
The Human Creativity Benchmark
Aspen Hopkins, Allison Nulty, Alexandria Minetti +2
Modern AI evaluation frameworks treat evaluator disagreement as noise to be resolved. In creative domains, professional disagreement reflects genuine differences in taste, not meas…
Large-Scale, Longitudinal Study of Large Language Models During the 2024 US Election Season
Sarah H. Cen, Andrew Ilyas, Hedi Driss +4
The 2024 US presidential election is the first major contest to occur in the US since the popularization of large language models (LLMs). Building on lessons from earlier shifts in…
Recourse, Repair, Reparation, & Prevention: A Stakeholder Analysis of AI Supply Chains
Aspen K. Hopkins, Isabella Struckman, Kevin Klyman +1
The AI industry is exploding in popularity, with increasing attention to potential harms and unwanted consequences. In the current digital ecosystem, AI deployments are often the p…
AI Supply Chains: An Emerging Ecosystem of AI Actors, Products, and Services
Aspen Hopkins, Sarah H. Cen, Andrew Ilyas +3
The widespread adoption of AI in recent years has led to the emergence of AI supply chains: complex networks of AI actors contributing models, datasets, and more to the development…