2 papers
cs.CR2026
Similarity Is Not Validity: Defending LLM Semantic Caches Against Poisoning
Zihan Zhang, Shuangjie Yao, Zesen Liu +8
Semantic caches reduce LLM serving costs by reusing previously generated answers for semantically similar queries. However, retrieval is based solely on embedding similarity betwee…
cs.CR2026
Rewriting the Response Path: Silent Tampering and Provider-Signed Defense in BYOK LLM Agents
Mingyu Luo, Zihan Zhang, Zesen Liu +7
LLM agents convert model outputs into consequential actions, including communications, code changes, and financial transactions. Developers often trust evidence such as test result…