1 citations · 1 across the 2 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2020
Fast WordPiece Tokenization
Xinying Song, Alex Salcianu, Yang Song +2
Tokenization is a fundamental preprocessing step for almost all NLP tasks. In this paper, we propose efficient algorithms for the WordPiece tokenization used in BERT, from single-w…
cs.CL2019
Extremely Small BERT Models from Mixed-Vocabulary Training
Sanqiang Zhao, Raghav Gupta, Yang Song +1
Pretrained language models like BERT have achieved good results on NLP tasks, but are impractical on resource-limited devices due to memory footprint. A large fraction of this foot…