2 citations · 3 across the 2 of their papers we have counts for
2 papers
cs.DC2024★ 2 cited
Lazarus: Resilient and Elastic Training of Mixture-of-Experts Models
Yongji Wu, Wenjie Qu, Xueshen Liu +10
Sparsely-activated Mixture-of-Experts (MoE) architecture has increasingly been adopted to further scale large language models (LLMs). However, frequent failures still pose signific…
cs.LG2020★ 1 cited
On InstaHide, Phase Retrieval, and Sparse Matrix Factorization
Sitan Chen, Xiaoxiao Li, Zhao Song +1
In this work, we examine the security of InstaHide, a scheme recently proposed by [Huang, Song, Li and Arora, ICML'20] for preserving the security of private datasets in the contex…