1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Hao Gu, Wei Li, Lujun Li +5
Mixture-of-Experts (MoE) architectures in large language models (LLMs) achieve exceptional performance, but face prohibitive storage and memory requirements. To address these chall…