563 citations · 1k across the 22 of their papers we have counts for
1 paper · 2 filters
Tianjun Zhang, Shishir G. Patil, Naman Jain +4
Pretraining Large Language Models (LLMs) on large corpora of textual data is now a standard paradigm. When using these LLMs for many downstream applications, it is common to additi…