1 paper
Seongryong Jung, Suwan Yoon, DongGeon Kim +1
Large language models (LLMs) offer impressive performance but are impractical for resource-constrained deployment due to high latency and energy consumption. Knowledge distillation…