1 paper
Jonathan Kamp, Roos Bakker, Dominique Blok
Good quality explanations strengthen the understanding of language models and data. Feature attribution methods, such as Integrated Gradient, are a type of post-hoc explainer that…