activity
20242026
most citedWho Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

1 citations · 1 across the 5 of their papers we have counts for

collaborators
Showing cs.CYShow all

5 papers · 1 filter

cs.CY2026

FLARE-AI: Flaw Reporting for AI

Shayne Longpre, Elaine Zhu, Carson Ezell +15

Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragmented: researchers who identify…

cs.CY20261 cited

Who Evaluates AI's Social Impacts? Mapping Coverage and Gaps in First and Third Party Evaluations

Anka Reuel, Avijit Ghosh, Jenny Chim +32

Foundation models are increasingly central to high-stakes AI systems, and governance frameworks now depend on evaluations to assess their risks and capabilities. Although general c…

cs.CY2026

What if AI systems weren't chatbots?

Sourojit Ghosh, Pranav Narayanan Venkit, Sanjana Gautam +1

The rapid convergence of artificial intelligence (AI) toward conversational chatbot interfaces marks a critical moment for the industry. This paper argues that the chatbot paradigm…

cs.CY2025

Economies of Open Intelligence: Tracing Power & Participation in the Model Ecosystem

Shayne Longpre, Christopher Akiki, Campbell Lund +6

Since 2019, the Hugging Face Model Hub has been the primary global platform for sharing open weight AI models. By releasing a dataset of the complete history of weekly model downlo…

cs.CY2024

To Err is AI : A Case Study Informing LLM Flaw Reporting Practices

Sean McGregor, Allyson Ettinger, Nick Judd +10

In August of 2024, 495 hackers generated evaluations in an open-ended bug bounty targeting the Open Language Model (OLMo) from The Allen Institute for AI. A vendor panel staffed by…