1 paper
Shangyu Liu, Zhenzhe Zheng, Xiaoyao Huang +3
Small language models (SLMs) support efficient deployments on resource-constrained edge devices, but their limited capacity compromises inference performance. Retrieval-augmented g…