Who Audits the Auditors? Recommendations from a field scan of the algorithmic auditing ecosystem
arXiv:2310.02521 · doi:10.1145/3531146.3533213
Abstract
AI audits are an increasingly popular mechanism for algorithmic accountability; however, they remain poorly defined. Without a clear understanding of audit practices, let alone widely used standards or regulatory guidance, claims that an AI product or system has been audited, whether by first-, second-, or third-party auditors, are difficult to verify and may exacerbate, rather than mitigate, bias and harm. To address this knowledge gap, we provide the first comprehensive field scan of the AI audit ecosystem. We share a catalog of individuals (N=438) and organizations (N=189) who engage in algorithmic audits or whose work is directly relevant to algorithmic audits; conduct an anonymous survey of the group (N=152); and interview industry leaders (N=10). We identify emerging best practices as well as methods and tools that are becoming commonplace, and enumerate common barriers to leveraging algorithmic audits as effective accountability mechanisms. We outline policy recommendations to improve the quality and impact of these audits, and highlight proposals with wide support from algorithmic auditors as well as areas of debate. Our recommendations have implications for lawmakers, regulators, internal company policymakers, and standards-setting bodies, as well as for auditors. They are: 1) require the owners and operators of AI systems to engage in independent algorithmic audits against clearly defined standards; 2) notify individuals when they are subject to algorithmic decision-making systems; 3) mandate disclosure of key components of audit findings for peer review; 4) consider real-world harm in the audit process, including through standardized harm incident reporting and response mechanisms; 5) directly involve the stakeholders most likely to be harmed by AI systems in the algorithmic audit process; and 6) formalize evaluation and, potentially, accreditation of algorithmic auditors.
20 pages, 2 figures. Published in the Proceedings of the 2022 ACM Conference on Fairness, Accountability, and Transparency (FAccT '22)
References in corpus (4)
- Improving fairness in machine learning systems: What do industry practitioners need?
- Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims
- Saving Face: Investigating the Ethical Concerns of Facial Recognition Auditing
- Algorithmic Auditing and Social Justice: Lessons from the History of Audit Studies
Cited by in corpus (23)
- Auditing large language models: a three-layered approach
- Auditing of AI: Legal, Ethical and Technical Approaches
- Walking the Walk of AI Ethics: Organizational Challenges and the Individualization of Risk among Ethics Entrepreneurs
- Black-Box Access is Insufficient for Rigorous AI Audits
- Certification Labels for Trustworthy AI: Insights From an Empirical Mixed-Method Study
- A Framework for Assurance Audits of Algorithmic Systems
- Sociotechnical Audits: Broadening the Algorithm Auditing Lens to Investigate Targeted Advertising
- Towards AI Accountability Infrastructure: Gaps and Opportunities in AI Audit Tooling
- How to Assess Trustworthy AI in Practice
- Participation and Division of Labor in User-Driven Algorithm Audits: How Do Everyday Users Work together to Surface Algorithmic Harms?
- Position is Power: System Prompts as a Mechanism of Bias in Large Language Models (LLMs)
- From human-centered to social-centered artificial intelligence: Assessing ChatGPT's impact through disruptive events
- Skin Deep: Investigating Subjectivity in Skin Tone Annotations for Computer Vision Benchmark Datasets
- (Beyond) Reasonable Doubt: Challenges that Public Defenders Face in Scrutinizing AI in Court
- Legacy Procurement Practices Shape How U.S. Cities Govern AI: Understanding Government Employees' Practices, Challenges, and Needs
- Access Denied: Meaningful Data Access for Quantitative Algorithm Audits
- Strategic Evaluation: Subjects, Evaluators, and Society
- "Unlimited Realm of Exploration and Experimentation": Methods and Motivations of AI-Generated Sexual Content Creators
- Algorithms in the Stacks: Investigating automated, for-profit diversity audits in public libraries
- Better Together? The Role of Explanations in Supporting Novices in Individual and Collective Deliberations about AI
- Interpretable and Fair Mechanisms for Abstaining Classifiers
- Compliant But Unsatisfactory: The Gap Between Auditing Standards and Practices for Probabilistic Genotyping Software
- AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence