1 paper
Daegun Yoon, Sangyoon Oh
To train deep learning models faster, distributed training on multiple GPUs is the very popular scheme in recent years. However, the communication bandwidth is still a major bottle…