1 paper
Luigi Altamura, Alessio Cicero, Mateo Vázquez Maceiras +3
The currently dominant AI/ML workloads, such as Large Language Models (LLMs), rely on the efficient execution of General Matrix-Matrix Multiplication (GEMM) operations. Thus, most…