1 paper
Aditya Agrawal, Albert Magyar, Hiteshwar Eswaraiah +5
Training and serving Large Language Models (LLMs) relies heavily on parallelization and collective operations, which are frequently bottlenecked by network bandwidth. Lossless comp…