2 papers
cs.CL2026
Friends and Grandmothers in Silico: Localizing Entity Cells in Language Models
Itay Yona, Dan Barzilay, Michael Karasik +1
How do language models retrieve entity-specific facts from their parameters? We investigate this question by searching for sparse, entity-selective MLP neurons - which we call enti…
cs.CL2025
In-Context Representation Hijacking
Itay Yona, Amir Sarid, Michael Karasik +1
We introduce , a simple in-context representation hijacking attack against large language models (LLMs). The attack works by systematically replacing a harmfu…