Dissociating language and thought in large language models
arXiv:2301.06627
Abstract
Large Language Models (LLMs) have come closest among all models to date to mastering human language, yet opinions about their linguistic and cognitive capabilities remain split. Here, we evaluate LLMs using a distinction between formal linguistic competence -- knowledge of linguistic rules and patterns -- and functional linguistic competence -- understanding and using language in the world. We ground this distinction in human neuroscience, which has shown that formal and functional competence rely on different neural mechanisms. Although LLMs are surprisingly good at formal competence, their performance on functional competence tasks remains spotty and often requires specialized fine-tuning and/or coupling with external modules. We posit that models that use language in human-like ways would need to master both of these competence types, which, in turn, could require the emergence of mechanisms specialized for formal linguistic competence, distinct from functional competence.
The two lead authors contributed equally to this work; published in "Trends in Cognnitive Sciences", March 2024
Cited by in corpus (11)
- Theory of Mind for Multi-Agent Collaboration via Large Language Models
- Exploring the Frontiers of LLMs in Psychological Applications: A Comprehensive Review
- Do LLMs Understand Social Knowledge? Evaluating the Sociability of Large Language Models with SocKET Benchmark
- Clinical Insights: A Comprehensive Review of Language Models in Medicine
- Computational Argumentation-based Chatbots: a Survey
- Programming-by-Demonstration for Long-Horizon Robot Tasks
- Can Language Models Be Tricked by Language Illusions? Easier with Syntax, Harder with Semantics
- Assessing the nature of large language models: A caution against anthropocentrism
- Large Language Models are biased to overestimate profoundness
- Mini Minds: Exploring Bebeshka and Zlata Baby Models
- Knowledge graphs for empirical concept retrieval