1 paper
Aashiq Muhamed, Oscar Li, David Woodruff +2
Large language model (LLM) training and finetuning are often bottlenecked by limited GPU memory. While existing projection-based optimization methods address this by projecting gra…