1 paper · 1 filter
Jianyu Wei, Shijie Cao, Ting Cao +4
The deployment of Large Language Models (LLMs) on edge devices is increasingly important to enhance on-device intelligence. Weight quantization is crucial for reducing the memory f…