16 citations · 16 across the 1 of their papers we have counts for
2 papers
cs.LG2018
Confidence Scoring Using Whitebox Meta-models with Linear Classifier Probes
Tongfei Chen, Jiří Navrátil, Vijay Iyengar +1
We propose a novel confidence scoring mechanism for deep neural networks based on a two-model paradigm involving a base model and a meta-model. The confidence score is learned by t…
cs.AI2017★ 16 cited
A Formal Framework to Characterize Interpretability of Procedures
Amit Dhurandhar, Vijay Iyengar, Ronny Luss +1
We provide a novel notion of what it means to be interpretable, looking past the usual association with human understanding. Our key insight is that interpretability is not an abso…