1 paper
Ceyu Xu, Yongji Wu, Xinyu Yang +4
As the parameter size of large language models (LLMs) continues to expand, the need for a large memory footprint and high communication bandwidth have become significant bottleneck…