978 citations · 1.5k across the 15 of their papers we have counts for
1 paper · 2 filters
Chenggang Zhao, Chengqi Deng, Chong Ruan +12
The rapid scaling of large language models (LLMs) has unveiled critical limitations in current hardware architectures, including constraints in memory capacity, computational effic…