47 citations · 72 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 25 cited
Post hoc Explanations may be Ineffective for Detecting Unknown Spurious Correlation
Julius Adebayo, Michael Muelly, Hal Abelson +1
We investigate whether three types of post hoc model explanations--feature attribution, concept activation, and training point ranking--are effective for detecting a model's relian…
cs.LG2016★ 47 cited
Iterative Orthogonal Feature Projection for Diagnosing Bias in Black-Box Models
Julius Adebayo, Lalana Kagal
Predictive models are increasingly deployed for the purpose of determining access to services such as credit, insurance, and employment. Despite potential gains in productivity and…