3 citations · 3 across the 3 of their papers we have counts for
1 paper · 1 filter
Saibo Geng, Sankalp Gambhir, Chris Wendler +1
Tokenization is an important preprocessing step in the training and inference of large language models (LLMs). While there has been extensive research on the expressive power of th…