1 paper
Lijia Liu, Takumi Kondo, Kyohei Atarashi +4
This paper investigates defenses for LLM-based evaluation systems against prompt injection. We formalize a class of threats called blind attacks, where a candidate answer is crafte…