1 paper · 1 filter
Linglin Jing, Yuting Gao, Zhigang Wang +5
Recent advancements have shown that the Mixture of Experts (MoE) approach significantly enhances the capacity of large language models (LLMs) and improves performance on downstream…