649 citations · 910 across the 9 of their papers we have counts for
3 papers · 1 filter
Bias Mitigation Framework for Intersectional Subgroups in Neural Networks
Narine Kokhlikyan, Bilal Alsallakh, Fulton Wang +3
We propose a fairness-aware learning framework that mitigates intersectional subgroup bias associated with protected attributes. Prior research has primarily focused on mitigating…
Investigating sanity checks for saliency maps with image and text classification
Narine Kokhlikyan, Vivek Miglani, Bilal Alsallakh +2
Saliency maps have shown to be both useful and misleading for explaining model predictions especially in the context of images. In this paper, we perform sanity checks for text mod…
Captum: A unified and generic model interpretability library for PyTorch
Narine Kokhlikyan, Vivek Miglani, Miguel Martin +8
In this paper we introduce a novel, unified, open-source model interpretability library for PyTorch [12]. The library contains generic implementations of a number of gradient and p…