3 papers
cs.CR2026
Trigger the Straggler: Load Hijack on Mixture-of-Experts LLMs
Rui Zhang, Wenbo Jiang, Hongwei Li +4
Expert parallelism (EP) is a common strategy for serving large Mixture-of-Experts (MoE) models across multiple GPUs by distributing experts among devices. Router decisions then det…
cs.CR2026
InkShield: Writing Style Protection Against Unauthorized Handwriting Mimicry
Jian Xiong, Wenbo Jiang, Zihan Wang +4
InkShield introduces a proactive defense that adds subtle, stroke‑confined perturbations to handwritten reference images, making it harder for handwriting generators to mimic a wri…
cs.CR2026
BadTemplate: A Training-Free Backdoor Attack via Chat Template Against Large Language Models
Zihan Wang, Hongwei Li, Rui Zhang +2
Chat template is a common technique used in the training and inference stages of Large Language Models (LLMs). It can transform input and output data into role-based and templated…