Showing cs.LGShow all
2 papers · 1 filter
cs.LG2026
The Assistant as a Privileged Persona: A canonical reference in cross-persona self-recognition
Asvin G
Post-trained language models can recognize their own outputs from a sentence or two out of context. In a companion paper \citep{jack2026twomodes} we showed they can also recognize…
cs.LG2026
From Simulation to Enaction: Post-trained language models recognize and react to their own generations
Asvin G., Jack Lindsey
Language models are pretrained as passive predictors with no incentive to model the consequences of their own outputs. Post-training changes this: a model producing its own respons…