1 paper · 1 filter
Jegwang Ryu, Minkyu Kim, Seungjun Shin +3
Efficient compression of language model weights is increasingly critical as model scale and deployment grow. Yet, most existing methods rely on handcrafted transforms and heuristic…