1 paper
Sixiao Huang, Tintin Wang, Ang Li +5
Large language models (LLMs) are both storage-intensive and computation-intensive, posing significant challenges when deployed on resource-constrained hardware. As linear layers in…