11 citations · 11 across the 4 of their papers we have counts for
3 papers · 1 filter
Confidential LLM Inference: Performance and Cost Across CPU and GPU TEEs
Marcin Chrapek, Marcin Copik, Etienne Mettaz +1
Large Language Models (LLMs) are increasingly deployed on converged Cloud and High-Performance Computing (HPC) infrastructure. However, as LLMs handle confidential inputs and are f…
Denoising Application Performance Models with Noise-Resilient Priors
Gustavo de Morais, Alexander GeiÃ, Alexandru Calotoiu +5
As parallel codes are scaled to larger computing systems, performance models play a crucial role in identifying potential bottlenecks. However, constructing these models analytical…
A Priori Loop Nest Normalization: Automatic Loop Scheduling in Complex Applications
Lukas Trümper, Philipp Schaad, Berke Ates +3
The same computations are often expressed differently across software projects and programming languages. In particular, how computations involving loops are expressed varies due t…