From the 1 of 7 linked papers with an AI index.
Showing cs.CLShow all
2 papers · 1 filter
cs.CL2026
Misalignment Has a Personality: A Big Five Account of Emergent Misalignment
Hasibur Rahman, Smit Desai
The paper proposes that misalignment in fine‑tuned language models can be understood as shifts in Big Five personality traits, extracting calibrated personality vectors that predic…
cs.CL2026
Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework
Hasibur Rahman, Smit Desai
Large language model (LLM)-based conversational agents (CAs) are now ubiquitous, creating new opportunities for AI-mediated behavior change. Their capacity to project nuanced perso…