1 citations · 1 across the 1 of their papers we have counts for
4 papers
How Causal Abstraction Underpins Computational Explanation
Atticus Geiger, Jacqueline Harding, Thomas Icard
Explanations of cognitive behavior often appeal to computations over representations. What does it take for a system to implement a given computation over suitable representational…
What Is AI Safety? What Do We Want It to Be?
Jacqueline Harding, Cameron Domenico Kirk-Giannini
The field of AI safety seeks to prevent or reduce the harms caused by AI systems. A simple and appealing account of what is distinctive of AI safety as a field holds that this feat…
A Communication-First Account of Explanation
Jacqueline Harding, Tobias Gerstenberg, Thomas Icard
This paper develops a formal account of causal explanation, grounded in a theory of conversational pragmatics, and inspired by the interventionist idea that explanation is about as…
What is it for a Machine Learning Model to Have a Capability?
Jacqueline Harding, Nathaniel Sharadin
What can contemporary machine learning (ML) models do? Given the proliferation of ML models in society, answering this question matters to a variety of stakeholders, both public an…