1 paper
Guihang Hong, Tao Ouyang, Kongyange Zhao +2
Motivated by the imperative for real-time responsiveness and data privacy preservation, large language models (LLMs) are increasingly deployed on resource-constrained edge devices…