1 paper
Long Shi, Bingyan Ou, Kang Wei +3
The sparse activation mechanism of mixture of experts (MoE) model empowers edge intelligence with enhanced training efficiency and reduced computational resource consumption. Howev…