1 paper
Jianwei Zhu, Hang Yin, Peng Deng +2
This report evaluates the performance impact of enabling Trusted Execution Environments (TEE) on NVIDIA Hopper GPUs for large language model (LLM) inference tasks. We benchmark the…