69 citations · 228 across the 16 of their papers we have counts for
3 papers · 2 filters
Preserving Causal Constraints in Counterfactual Explanations for Machine Learning Classifiers
Divyat Mahajan, Chenhao Tan, Amit Sharma
To construct interpretable explanations that are consistent with the original ML model, counterfactual examples---showing how the model's output changes with small perturbations to…
Explaining Machine Learning Classifiers through Diverse Counterfactual Explanations
Ramaravind Kommiya Mothilal, Amit Sharma, Chenhao Tan
Post-hoc explanations of machine learning models are crucial for people to understand and act on algorithmic predictions. An intriguing class of explanations is through counterfact…
Learning Fair Representations via an Adversarial Framework
Rui Feng, Yang Yang, Yuehan Lyu +3
Fairness has become a central issue for our research community as classification algorithms are adopted in societally critical domains such as recidivism prediction and loan approv…