16 citations · 16 across the 4 of their papers we have counts for
5 papers
A Roadmap for Big Model
Sha Yuan, Hanyu Zhao, Shuai Zhao +97
With the rapid development of deep learning, training Big Models (BMs) for multiple downstream tasks becomes a popular paradigm. Researchers have achieved various outcomes in the c…
A New Adaptive Gradient Method with Gradient Decomposition
Zhou Shao, Tong Lin
Adaptive gradient methods, especially Adam-type methods (such as Adam, AMSGrad, and AdaBound), have been proposed to speed up the training process with an element-wise scaling term…
Turing Award elites revisited: patterns of productivity, collaboration, authorship and impact
Yinyu Jin, Sha Yuan, Zhou Shao +2
The Turing Award is recognized as the most influential and prestigious award in the field of computer science(CS). With the rise of the science of science (SciSci), a large amount…
CogView: Mastering Text-to-Image Generation via Transformers
Ming Ding, Zhuoyi Yang, Wenyi Hong +8
Text-to-Image generation in the general domain has long been an open problem, which requires both a powerful generative model and cross-modal understanding. We propose CogView, a 4…
Attention: to Better Stand on the Shoulders of Giants
Sha Yuan, Zhou Shao, Yu Zhang +4
Science of science (SciSci) is an emerging discipline wherein science is used to study the structure and evolution of science itself using large data sets. The increasing availabil…