3 papers
cs.DC2024
Accelerating Distributed Deep Learning using Lossless Homomorphic Compression
Haoyu Li, Yuchen Xu, Jiayi Chen +5
As deep neural networks (DNNs) grow in complexity and size, the resultant increase in communication overhead during distributed training has become a significant bottleneck, challe…
cs.CL2023
Mitigating the Exposure Bias in Sentence-Level Grapheme-to-Phoneme (G2P) Transduction
Eunseop Yoon, Hee Suk Yoon, Dhananjaya Gowda +7
Text-to-Text Transfer Transformer (T5) has recently been considered for the Grapheme-to-Phoneme (G2P) transduction. As a follow-up, a tokenizer-free byte-level model based on T5 re…
cs.CV2022
Self-Supervised Visual Representation Learning via Residual Momentum
Trung X. Pham, Axi Niu, Zhang Kang +5
Self-supervised learning (SSL) approaches have shown promising capabilities in learning the representation from unlabeled data. Amongst them, momentum-based frameworks have attract…