15 citations · 19 across the 2 of their papers we have counts for
2 papers
cs.DC2023★ 15 cited
ZeRO++: Extremely Efficient Collective Communication for Giant Model Training
Guanhua Wang, Heyang Qin, Sam Ade Jacobs +6
Zero Redundancy Optimizer (ZeRO) has been used to train a wide range of large language models on massive GPUs clusters due to its ease of use, efficiency, and good scalability. How…
cs.CV2023★ 4 cited
LayoutDM: Transformer-based Diffusion Model for Layout Generation
Shang Chai, Liansheng Zhuang, Fengying Yan
Automatic layout generation that can synthesize high-quality layouts is an important tool for graphic design in many applications. Though existing methods based on generative model…