2 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Ryan Solgi, Kai Zhen, Rupak Vignesh Swaminathan +4
The efficient implementation of large language models (LLMs) is crucial for deployment on resource-constrained devices. Low-rank tensor compression techniques, such as tensor-train…