1 paper · 1 filter
Jun Hou, Le Wang, Xuan Wang
Mixture-of-Experts (MoE) models have become increasingly powerful in multimodal learning by enabling modular specialization across modalities. However, their effectiveness remains…