Showing cs.CRShow all
2 papers · 1 filter
cs.CR2026
Here is a GIFT: Enforcing User Data Isolation in LLM Serving via GPU Information Flow Tracking
Jiacheng Shi, Xunjie Wang, Cheng Tan +1
LLM serving frameworks process large volumes of user data--often containing sensitive information--on shared infrastructure. Ensuring isolation between users who share the same ser…
cs.CR2024
PipeLLM: Fast and Confidential Large Language Model Services with Speculative Pipelined Encryption
Yifan Tan, Cheng Tan, Zeyu Mi +1
Confidential computing on GPUs, like NVIDIA H100, mitigates the security risks of outsourced Large Language Models (LLMs) by implementing strong isolation and data encryption. None…