25 citations · 25 across the 3 of their papers we have counts for
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2023
YUAN 2.0: A Large Language Model with Localized Filtering-based Attention
Shaohua Wu, Xudong Zhao, Shenling Wang +9
In this work, we develop and release Yuan 2.0, a series of large language models with parameters ranging from 2.1 billion to 102.6 billion. The Localized Filtering-based Attention…
cs.CL2021★ 25 cited
Yuan 1.0: Large-Scale Pre-trained Language Model in Zero-Shot and Few-Shot Learning
Shaohua Wu, Xudong Zhao, Tong Yu +8
Recent work like GPT-3 has demonstrated excellent performance of Zero-Shot and Few-Shot learning on many natural language processing (NLP) tasks by scaling up model size, dataset s…