25 citations · 32 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 25 cited
Post hoc Explanations may be Ineffective for Detecting Unknown Spurious Correlation
Julius Adebayo, Michael Muelly, Hal Abelson +1
We investigate whether three types of post hoc model explanations--feature attribution, concept activation, and training point ranking--are effective for detecting a model's relian…
cs.CV2020★ 7 cited
Debugging Tests for Model Explanations
Julius Adebayo, Michael Muelly, Ilaria Liccardi +1
We investigate whether post-hoc model explanations are effective for diagnosing model errors--model debugging. In response to the challenge of explaining a model's prediction, a va…