1 paper
Fangxin Liu, Zongwu Wang, JinHong Xia +6
The rapid advancement of large language models (LLMs) has exacerbated the memory bottleneck due to the widening gap between model parameter scaling and hardware capabilities. While…