1 paper · 1 filter
Shaohua Wu, Xudong Zhao, Shenling Wang +9
In this work, we develop and release Yuan 2.0, a series of large language models with parameters ranging from 2.1 billion to 102.6 billion. The Localized Filtering-based Attention…