35 citations · 35 across the 1 of their papers we have counts for
2 papers
cs.LG2020★ 35 cited
Fairwashing Explanations with Off-Manifold Detergent
Christopher J. Anders, Plamen Pasliev, Ann-Kathrin Dombrowski +2
Explanation methods promise to make black-box classifiers more transparent. As a result, it is hoped that they can act as proof for a sensible, fair and trustworthy decision-making…
stat.ML2019
Explanations can be manipulated and geometry is to blame
Ann-Kathrin Dombrowski, Maximilian Alber, Christopher J. Anders +3
Explanation methods aim to make neural networks more trustworthy and interpretable. In this paper, we demonstrate a property of explanation methods which is disconcerting for both…