1 paper
Ahin Lee, Sehyun Yun, Taesik Gong
Mixture-of-Experts (MoE) models scale efficiently but remain costly to adapt due to redundant experts and uniform parameter allocation. Existing parameter-efficient fine-tuning (PE…