16 citations · 26 across the 13 of their papers we have counts for
Showing 2020 · cs.CLShow all
2 papers · 2 filters
cs.CL2020
DICT-MLM: Improved Multilingual Pre-Training using Bilingual Dictionaries
Aditi Chaudhary, Karthik Raman, Krishna Srinivasan +1
Pre-trained multilingual language models such as mBERT have shown immense gains for several natural language processing (NLP) tasks, especially in the zero-shot cross-lingual setti…
cs.CL2020
DiPair: Fast and Accurate Distillation for Trillion-Scale Text Matching and Pair Modeling
Jiecao Chen, Liu Yang, Karthik Raman +6
Pre-trained models like BERT (Devlin et al., 2018) have dominated NLP / IR applications such as single sentence classification, text pair classification, and question answering. Ho…