Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
The Bridge-Garden Dilemma in LLM Distillation: Why Mixing Hard and Soft Labels Works
Guanghui Wang, Kaiwen Lv Kacuila, Zhiyong Yang +5
Knowledge distillation (KD) transfers knowledge from a large teacher model to a smaller student. In language modeling, the student is trained either on tokens sampled from the teac…
cs.LG2026
Localize and Neutralize: Gradient-guided Token Suppression against Visual Prompt Injection Attack
Dongpeng Zhang, Ke Ma, Yangbangyan Jiang +4
Adversarial images pose a severe security threat to multimodal large language models through prompt injection. Existing defenses largely lack a principled understanding of the unde…