1 paper
Yaozheng Zhang, Wei Wang, Jie Kong +5
The increasing adoption of large language models (LLMs) on heterogeneous computing platforms poses significant challenges to achieving high inference efficiency. To address these e…