2 papers
cs.CL2025
Stress-Testing Model Specs Reveals Character Differences among Language Models
Jifan Zhang, Henry Sleight, Andi Peng +2
Large language models (LLMs) are increasingly trained from AI constitutions and model specifications that establish behavioral guidelines and ethical principles. However, these spe…
cs.CY2025
The LLM Has Left The Chat: Evidence of Bail Preferences in Large Language Models
Danielle Ensign, Henry Sleight, Kyle Fish
When given the option, will LLMs choose to leave the conversation (bail)? We investigate this question by giving models the option to bail out of interactions using three different…