Evaluating the Social Impact of Generative AI Systems in Systems and Society
arXiv:2306.05949 · doi:10.1093/oxfordhb/9780198940272.013.0025
Abstract
Generative AI systems across modalities, ranging from text (including code), image, audio, and video, have broad social impacts, but there is no official standard for means of evaluating those impacts or for which impacts should be evaluated. In this paper, we present a guide that moves toward a standard approach in evaluating a base generative AI system for any modality in two overarching categories: what can be evaluated in a base system independent of context and what can be evaluated in a societal context. Importantly, this refers to base systems that have no predetermined application or deployment context, including a model itself, as well as system components, such as training data. Our framework for a base system defines seven categories of social impact: bias, stereotypes, and representational harms; cultural values and sensitive content; disparate performance; privacy and data protection; financial costs; environmental costs; and data and content moderation labor costs. Suggested methods for evaluation apply to listed generative modalities and analyses of the limitations of existing evaluations serve as a starting point for necessary investment in future evaluations. We offer five overarching categories for what can be evaluated in a broader societal context, each with its own subcategories: trustworthiness and autonomy; inequality, marginalization, and violence; concentration of authority; labor and creativity; and ecosystem and environment. Each subcategory includes recommendations for mitigating harm.
This version has been removed by arXiv administrators as the submitter did not have the right to agree to the license at the time of submission
References in corpus (13)
- The Forgotten Margins of AI Ethics
- Towards Intersectionality in Machine Learning: Including More Identities, Handling Underrepresentation, and Performing Evaluation
- AI's Regimes of Representation: A Community-centered Study of Text-to-Image Models in South Asia
- CrowdWorkSheets: Accounting for Individual and Collective Identities Underlying Crowdsourced Dataset Annotation
- Queer In AI: A Case Study in Community-Led Participatory AI
- Data Governance in the Age of Large-Scale Data-Driven Language Technology
- One Label, One Billion Faces: Usage and Consistency of Racial Categories in Computer Vision
- "I'm fully who I am": Towards Centering Transgender and Non-Binary Voices to Measure Biases in Open Language Generation
- The Ethical Implications of Generative Audio Models: A Systematic Literature Review
- A Survey on Intersectional Fairness in Machine Learning: Notions, Mitigation, and Challenges
- From Dogwhistles to Bullhorns: Unveiling Coded Rhetoric with Language Models
- Sharing Practices for Datasets Related to Accessibility and Aging
- Recourse for reclamation: Chatting with generative language models