1 paper · 1 filter
Jehyeon Bang, Eunyeong Cho, Ranggi Hwang +2
The Mixture-of-Experts (MoE) architecture has emerged as a promising approach to mitigate the rising computational costs of large language models (LLMs) by selectively activating p…