1 paper
HamidReza Imani, Jiaxin Peng, Peiman Mohseni +2
The deployment of mixture-of-experts (MoE) large language models (LLMs) presents significant challenges due to their high memory demands. These challenges become even more pronounc…