19 citations · 21 across the 8 of their papers we have counts for
Showing cs.CLShow all
3 papers · 1 filter
cs.CL2020
Inserting Information Bottlenecks for Attribution in Transformers
Zhiying Jiang, Raphael Tang, Ji Xin +1
Pretrained transformers achieve the state of the art across tasks in natural language processing, motivating researchers to investigate their inner mechanisms. One common direction…
cs.CL2020★ 1 cited
Showing Your Work Doesn't Always Work
Raphael Tang, Jaejun Lee, Ji Xin +3
In natural language processing, a recently popular line of work explores how to best report the experimental results of neural networks. One exemplar publication, titled "Show Your…
cs.CL2020★ 19 cited
DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference
Ji Xin, Raphael Tang, Jaejun Lee +2
Large-scale pre-trained language models such as BERT have brought significant improvements to NLP applications. However, they are also notorious for being slow in inference, which…