1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.DC2025
MoEntwine: Unleashing the Potential of Wafer-scale Chips for Large-scale Expert Parallel Inference
Xinru Tang, Jingxiang Hou, Dingcheng Jiang +9
As large language models (LLMs) continue to scale up, mixture-of-experts (MoE) has become a common technology in SOTA models. MoE models rely on expert parallelism (EP) to alleviat…
cs.AR2023★ 1 cited
Wafer-scale Computing: Advancements, Challenges, and Future Perspectives
Yang Hu, Xinhan Lin, Huizheng Wang +12
Nowadays, artificial intelligence (AI) technology with large models plays an increasingly important role in both academia and industry. It also brings a rapidly increasing demand f…