1 paper
Yibin Luo, Shiwei Gao, Huichuan Zheng +2
Fine-tuning Large Language Models (LLMs) on consumer-grade GPUs is highly cost-effective, yet constrained by limited GPU memory and slow PCIe interconnects. Pipeline parallelism co…