1 citations · 2 across the 2 of their papers we have counts for
Showing 2020Show all
2 papers · 1 filter
cs.CL2020
Causal Mediation Analysis for Interpreting Neural NLP: The Case of Gender Bias
Jesse Vig, Sebastian Gehrmann, Yonatan Belinkov +6
Common methods for interpreting neural models in natural language processing typically examine either their structure or their behavior, but not both. We propose a methodology grou…
cs.LG2020★ 1 cited
Robustness from Simple Classifiers
Sharon Qian, Dimitris Kalimeris, Gal Kaplun +1
Despite the vast success of Deep Neural Networks in numerous application domains, it has been shown that such models are not robust i.e., they are vulnerable to small adversarial p…