2 citations · 2 across the 1 of their papers we have counts for
1 paper
Jonathan Helland, Nathan VanHoudnos
In this work, we investigate the phenomenon that robust image classifiers have human-recognizable features -- often referred to as interpretability -- as revealed through the input…