17 citations · 20 across the 2 of their papers we have counts for
2 papers
cs.CL2026★ 3 cited
The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models
Christina Lu, Jack Gallagher, Jonathan Michala +2
Large language models can represent a variety of personas but typically default to a helpful Assistant identity cultivated during post-training. We investigate the structure of the…
cs.LG2022★ 17 cited
Subverting machines, fluctuating identities: Re-learning human categorization
Christina Lu, Jackie Kay, Kevin R. McKee
Most machine learning systems that interact with humans construct some notion of a person's "identity," yet the default paradigm in AI research envisions identity with essential at…