31 citations · 31 across the 2 of their papers we have counts for
2 papers
cs.LG2026
LightMoE: Reducing Mixture-of-Experts Redundancy through Expert Replacing
Jiawei Hao, Zhiwei Hao, Jianyuan Guo +4
Mixture-of-Experts (MoE) based Large Language Models (LLMs) have demonstrated impressive performance and computational efficiency. However, their deployment is often constrained by…
cs.CV2023★ 31 cited
LGViT: Dynamic Early Exiting for Accelerating Vision Transformer
Guanyu Xu, Jiawei Hao, Li Shen +4
Recently, the efficient deployment and acceleration of powerful vision transformers (ViTs) on resource-limited edge devices for providing multimedia services have become attractive…