1 paper
Felix Schneider, Maria Gogolev, Sven Sickert +1
Tokenization and sub-tokenization based models like word2vec, BERT and the GPTs are the state-of-the-art in natural language processing. Typically, these approaches have limitation…