1 paper · 1 filter
Mingyu Ma, Yuxin Wu, Jingbo Wang +3
Large Language Models (LLMs) are increasingly deployed with hierarchical instructions, yet they remain vulnerable to conflicts in which user directives override system-level constr…