1 paper · 1 filter
Shien Zhu, Samuel Bohl, Robin Oester +1
Mixture-of-Experts (MoE) Large Language Models (LLMs) efficiently scale-up the model while keeping relatively low inference cost. As MoE models only activate part of the experts, r…