Six misconceptions about large language models: A minimal model and diagnostic taxonomy
arXiv:2608.20421 · doi:10.1093/pnasnexus/pgag236
Abstract
Large language models (LLMs) are now embedded in scientific, educational, and governance workflows, with debates centering on their capabilities, mechanisms, and impacts. Yet these debates remain structured by persistent folk theories--intuitive, informal explanatory models that guide attitudes and actions. Deflationary slogans ("just autocomplete," "stochastic parrots," and "average of the internet") and anthropomorphic framings ("emergent agents" and "proto-minds") each capture genuine features of current systems but mistake those features for the whole. This Perspective proposes a minimal working model of LLM-based systems centered on four distinctions: between pretraining and deployed systems; between the learned distribution and particular samples; among parametric, contextual, and external memory; and between task competence and agency. The model is used to diagnose six misconceptions about LLMs: next-token prediction, regression to the mean, training-data regurgitation, model memory, alignment, and understanding. For each, the analysis identifies what the misconception gets right, which distinctions it conflates, and what follows for capability evaluation, system design, and governance. Applied to publisher AI policies as governance case studies, the framework shows both how policy language can conflate these distinctions and how such errors can be corrected. The model thereby avoids the parrot-mind binary by treating LLMs as simulators of discourse and task performance, offering a diagnostic toolkit for locating and correcting the errors these folk theories perpetuate.
20 pages, 1 figure, 2 tables, and 2 boxes. Published in PNAS Nexus
References in corpus (15)
- On the Opportunities and Risks of Foundation Models
- A Survey on Large Language Model based Autonomous Agents
- Sparks of Artificial General Intelligence: Early experiments with GPT-4
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning
- Generative artificial intelligence enhances creativity but reduces the diversity of novel content
- Fine-Tuning Language Models from Human Preferences
- MusicLM: Generating Music From Text
- GeneGPT: Augmenting Large Language Models with Domain Tools for Improved Access to Biomedical Information
- Language agents achieve superhuman synthesis of scientific knowledge
- How ChatGPT Changed the Media's Narratives on AI: A Semi-Automated Narrative Analysis Through Frame Semantics
- Superhuman performance of a large language model on the reasoning tasks of a physician
- From tools to thieves: Measuring and understanding public perceptions of AI through crowdsourced metaphors
- Mathematical exploration and discovery at scale
- From Human Memory to AI Memory: A Survey on Memory Mechanisms in the Era of LLMs
- Memory in Large Language Models: Mechanisms, Evaluation and Evolution