25 citations · 32 across the 2 of their papers we have counts for
3 papers
Post hoc Explanations may be Ineffective for Detecting Unknown Spurious Correlation
Julius Adebayo, Michael Muelly, Hal Abelson +1
We investigate whether three types of post hoc model explanations--feature attribution, concept activation, and training point ranking--are effective for detecting a model's relian…
Debugging Tests for Model Explanations
Julius Adebayo, Michael Muelly, Ilaria Liccardi +1
We investigate whether post-hoc model explanations are effective for diagnosing model errors--model debugging. In response to the challenge of explaining a model's prediction, a va…
Sanity Checks for Saliency Maps
Julius Adebayo, Justin Gilmer, Michael Muelly +3
Saliency methods have emerged as a popular tool to highlight features in an input deemed relevant for the prediction of a learned model. Several saliency methods have been proposed…