4 papers
Do as I Say, Not as I Do: Instruction-Induction Conflict in LLMs
Carolina Camassa, Derek Shiller
Language models are trained to follow instructions, but they are also powerful pattern completers. What happens when these two objectives conflict? We construct conversations in wh…
Initial results of the Digital Consciousness Model
Derek Shiller, Laura Duffy, Arvo Muñoz Morán +3
Artificially intelligent systems have become remarkably sophisticated. They hold conversations, write essays, and seem to understand context in ways that surprise even their creato…
Consciousness in Artificial Intelligence? A Framework for Classifying Objections and Constraints
Andres Campero, Derek Shiller, Jaan Aru +1
We develop a taxonomical framework for classifying challenges to the possibility of consciousness in digital artificial intelligence systems. This framework allows us to identify t…
Beyond Mimicry: Preference Coherence in LLMs
Luhan Mikaelson, Derek Shiller, Hayley Clatterbuck
We investigate whether large language models exhibit genuine preference structures by testing their responses to AI-specific trade-offs involving GPU reduction, capability restrict…