57 citations · 170 across the 18 of their papers we have counts for
Showing 2022Show all
2 papers · 1 filter
cs.LG2022★ 57 cited
Galvatron: Efficient Transformer Training over Multiple GPUs Using Automatic Parallelism
Xupeng Miao, Yujie Wang, Youhe Jiang +4
Transformer models have achieved state-of-the-art performance on various domains of applications and gradually becomes the foundations of the advanced large deep learning (DL) mode…
cs.CV2022★ 10 cited
Diffusion-Based Scene Graph to Image Generation with Masked Contrastive Pre-Training
Ling Yang, Zhilin Huang, Yang Song +6
Generating images from graph-structured inputs, such as scene graphs, is uniquely challenging due to the difficulty of aligning nodes and connections in graphs with objects and the…