1 paper
Yuanteng Chen, Qiwei Lai, Chen Tianqi +7
Fine-grained mixture-of-experts (MoE) architectures have become a mainstream design for open-weight LLMs, with hundreds of experts and increasingly many selected per token. This sh…