8 citations · 8 across the 2 of their papers we have counts for
2 papers
cs.CL2025
Pangu Ultra MoE: How to Train Your Big MoE on Ascend NPUs
Yehui Tang, Yichun Yin, Yaoyuan Wang +71
Sparse large language models (LLMs) with Mixture of Experts (MoE) and close to a trillion parameters are dominating the realm of most capable language models. However, the massive…
cs.AR2013★ 8 cited
MIMS: Towards a Message Interface based Memory System
Licheng Chen, Tianyue Lu, Yanan Wang +7
Memory system is often the main bottleneck in chipmultiprocessor (CMP) systems in terms of latency, bandwidth and efficiency, and recently additionally facing capacity and power pr…