1 paper
Ruifan Chu, Anbang Wang, Xiuxiu Bai +2
In high-performance computing, hotspot GPU kernels are primary bottlenecks, and expert manual tuning is costly and hard to port. Large language model methods often assume kernels c…