1 paper
Huakang Lin, Tiancheng Zheng, Mingxuan Sun +4
Mixture-of-Experts (MoE) architectures enable scalable and efficient large language models (LLMs) by selectively activating expert sub-networks through a routing mechanism. However…