128 citations · 130 across the 2 of their papers we have counts for
3 papers
AI and the Future of Digital Public Squares
Beth Goldberg, Diana Acosta-Navas, Michiel Bakker +24
Two substantial technological advances have reshaped the public square in recent decades: first with the advent of the internet and second with the recent introduction of large lan…
When Benchmarks are Targets: Revealing the Sensitivity of Large Language Model Leaderboards
Norah Alzahrani, Hisham Abdullah Alyahya, Yazeed Alnumay +9
Large Language Model (LLM) leaderboards based on benchmark rankings are regularly used to guide practitioners in model selection. Often, the published leaderboard rankings are take…
Holistic Evaluation of Language Models
Percy Liang, Rishi Bommasani, Tony Lee +47
Language models (LMs) are becoming the foundation for almost all major language technologies, but their capabilities, limitations, and risks are not well understood. We present Hol…