2 citations · 2 across the 8 of their papers we have counts for
1 paper · 1 filter
Qian Chen, Xianhao Chen, Kaibin Huang
Mixture-of-Experts (MoE) architectures leverage sparse activation to enhance the scalability of large language models (LLMs), making them suitable for deployment in resource-constr…