1 citations · 1 across the 7 of their papers we have counts for
1 paper · 1 filter
Haozheng Fan, Hao Zhou, Guangtai Huang +6
Getting large language models (LLMs) to perform well on the downstream tasks requires pre-training over trillions of tokens. This typically demands a large number of powerful compu…