3 citations · 5 across the 2 of their papers we have counts for
2 papers
cs.LG2024★ 3 cited
Harder Tasks Need More Experts: Dynamic Routing in MoE Models
Quzhe Huang, Zhenwei An, Nan Zhuang +7
In this paper, we introduce a novel dynamic expert selection framework for Mixture of Experts (MoE) models, aiming to enhance computational efficiency and model performance by adju…
cs.CL2023★ 2 cited
A Step Closer to Comprehensive Answers: Constrained Multi-Stage Question Decomposition with Large Language Models
Hejing Cao, Zhenwei An, Jiazhan Feng +3
While large language models exhibit remarkable performance in the Question Answering task, they are susceptible to hallucinations. Challenges arise when these models grapple with u…