1 paper
Oguzhan Baser, Elahe Sadeghi, Eric Wang +5
Most large language models (LLMs) run on external clouds: users send a prompt, pay for inference, and must trust that the remote GPU executes the LLM without any adversarial tamper…