1 paper
Yijun Lu, Zihan Fang, Pengpeng Qiao +6
The continuous scaling of large language models (LLMs) incurs prohibitive computational costs, making Mixture-of-Experts (MoE) a scalable alternative for efficient fine-tuning via…