10 citations · 16 across the 16 of their papers we have counts for
1 paper · 1 filter
Tongtian Yue, Longteng Guo, Jie Cheng +2
In the era of Large Language Models (LLMs), Mixture-of-Experts (MoE) architectures offer a promising approach to managing computational costs while scaling up model parameters. Con…