1 paper · 1 filter
Jinqi Xiao, Shen Sang, Tiancheng Zhi +5
Training large-scale neural networks in vision, and multimodal domains demands substantial memory resources, primarily due to the storage of optimizer states. While LoRA, a popular…