11 citations · 11 across the 3 of their papers we have counts for
Showing cs.PFShow all
2 papers · 1 filter
cs.PF2025
Confidential LLM Inference: Performance and Cost Across CPU and GPU TEEs
Marcin Chrapek, Marcin Copik, Etienne Mettaz +1
Large Language Models (LLMs) are increasingly deployed on converged Cloud and High-Performance Computing (HPC) infrastructure. However, as LLMs handle confidential inputs and are f…
cs.PF2024
A Priori Loop Nest Normalization: Automatic Loop Scheduling in Complex Applications
Lukas Trümper, Philipp Schaad, Berke Ates +3
The same computations are often expressed differently across software projects and programming languages. In particular, how computations involving loops are expressed varies due t…