2 papers
cs.CL2026
When Speculation Spills Secrets: Side Channels via Speculative Decoding In LLMs
Jiankun Wei, Abdulrahman Abdulrazzag, Tianchen Zhang +2
Deployed large language models (LLMs) often rely on speculative decoding, a technique that generates and verifies multiple candidate tokens in parallel, to improve throughput and l…
cs.LG2024
Time Will Tell: Timing Side Channels via Output Token Count in Large Language Models
Tianchen Zhang, Gururaj Saileshwar, David Lie
This paper demonstrates a new side-channel that enables an adversary to extract sensitive information about inference inputs in large language models (LLMs) based on the number of…