329 citations · 331 across the 3 of their papers we have counts for
3 papers
cs.LG2024★ 2 cited
Efficient Expert Pruning for Sparse Mixture-of-Experts Language Models: Enhancing Performance and Reducing Inference Costs
Enshu Liu, Junyi Zhu, Zinan Lin +6
The rapid advancement of large language models (LLMs) has led to architectures with billions to trillions of parameters, posing significant deployment challenges due to their subst…
cs.DC2023
Proteus: Simulating the Performance of Distributed DNN Training
Jiangfei Duan, Xiuhong Li, Ping Xu +4
DNN models are becoming increasingly larger to achieve unprecedented accuracy, and the accompanying increased computation and memory requirements necessitate the employment of mass…
cs.CV2015★ 329 cited
Deep Image: Scaling up Image Recognition
Ren Wu, Shengen Yan, Yi Shan +2
We present a state-of-the-art image recognition system, Deep Image, developed using end-to-end deep learning. The key components are a custom-built supercomputer dedicated to deep…