1 paper · 1 filter
Liujianfu Wang, Yuyang Du, Yuchen Pan +3
Mixture-of-Experts (MoE), while offering significant advantages as a Large Language Model (LLM) architecture, faces substantial challenges when deployed on low-cost edge devices wi…