2 citations · 2 across the 1 of their papers we have counts for
1 paper
Hongyu Wang, Shuming Ma, Ruiping Wang +1
We introduce, Q-Sparse, a simple yet effective approach to training sparsely-activated large language models (LLMs). Q-Sparse enables full sparsity of activations in LLMs which can…