3 papers
cs.DC2026
HyperOffload: Graph-Driven Hierarchical Memory Management for Large Language Models on SuperNode Architectures
Fangxin Liu, Qinghua Zhang, Hanjing Shen +5
The rapid evolution of Large Language Models (LLMs) towards long-context reasoning and sparse architectures has pushed memory requirements far beyond the capacity of individual dev…
cs.AI2026
SIEVE: Selective Integrity Verification and Escalation for Defending LLM Agents against Indirect Prompt Injection
Zhibo Liang, Tianze Hu, Zaiye Chen +1
Large Language Models (LLMs) are increasingly used as the core of agentic systems due to their strong reasoning, planning, and tool-use capabilities. By interacting with external e…
cs.SD2025
CloneShield: A Framework for Universal Perturbation Against Zero-Shot Voice Cloning
Renyuan Li, Zhibo Liang, Haichuan Zhang +5
Recent breakthroughs in text-to-speech (TTS) voice cloning have raised serious privacy concerns, allowing highly accurate vocal identity replication from just a few seconds of refe…