22 citations · 36 across the 2 of their papers we have counts for
2 papers
cs.LG2022★ 22 cited
Interpretability, Then What? Editing Machine Learning Models to Reflect Human Knowledge and Values
Zijie J. Wang, Alex Kale, Harsha Nori +6
Machine learning (ML) interpretability techniques can reveal undesirable patterns in data that models exploit to make predictions--potentially causing harms once deployed. However,…
cs.LG2021★ 14 cited
GAM Changer: Editing Generalized Additive Models with Interactive Visualization
Zijie J. Wang, Alex Kale, Harsha Nori +6
Recent strides in interpretable machine learning (ML) research reveal that models exploit undesirable patterns in the data to make predictions, which potentially causes harms in de…