1 citations · 1 across the 5 of their papers we have counts for
Showing cs.LGShow all
2 papers · 1 filter
cs.LG2025★ 1 cited
Mixture of Experts in Large Language Models
Danyang Zhang, Junhao Song, Ziqian Bi +5
This paper presents a comprehensive review of the Mixture-of-Experts (MoE) architecture in large language models, highlighting its ability to significantly enhance model performanc…
cs.LG2025
Multimodal Representation Learning and Fusion
Qihang Jin, Enze Ge, Yuhang Xie +8
Multi-modal learning is a fast growing area in artificial intelligence. It tries to help machines understand complex things by combining information from different sources, like im…