1 paper
Sifeng Shang, Jiayi Zhou, Chenyu Lin +2
As the size of large language models grows exponentially, GPU memory has become a bottleneck for adapting these models to downstream tasks. In this paper, we aim to push the limits…