11 citations · 11 across the 1 of their papers we have counts for
2 papers
cs.CL2021★ 11 cited
An Interpretability Illusion for BERT
Tolga Bolukbasi, Adam Pearce, Ann Yuan +4
We describe an "interpretability illusion" that arises when analyzing the BERT model. Activations of individual neurons in the network may spuriously appear to encode a single, sim…
cs.LG2019
Visualizing and Measuring the Geometry of BERT
Andy Coenen, Emily Reif, Ann Yuan +4
Transformer architectures show significant promise for natural language processing. Given that a single pretrained model can be fine-tuned to perform well on many different tasks,…