1 paper
Gautom Das, Vincent La, Ethan Lau +2
Large language models (LLMs) deliver impressive results for a variety of tasks, but state-of-the-art systems require fast GPUs with large amounts of memory. To reduce both the memo…