Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
The Assistant Axis: Situating and Stabilizing the Default Persona of Language Models
Christina Lu, Jack Gallagher, Jonathan Michala +2
Large language models can represent a variety of personas but typically default to a helpful Assistant identity cultivated during post-training. We investigate the structure of the…
cs.CL2025
Mechanistic Decomposition of Sentence Representations
Matthieu Tehenan, Vikram Natarajan, Jonathan Michala +2
Sentence embeddings are central to modern NLP and AI systems, yet little is known about their internal structure. While we can compare these embeddings using measures such as cosin…