From Bytes to Biases: Investigating the Cultural Self-Perception of Large Language Models
arXiv:2312.17256 · doi:10.1177/07439156251319788
Abstract
Large language models (LLMs) are able to engage in natural-sounding conversations with humans, showcasing unprecedented capabilities for information retrieval and automated decision support. They have disrupted human-technology interaction and the way businesses operate. However, technologies based on generative artificial intelligence (GenAI) are known to hallucinate, misinform, and display biases introduced by the massive datasets on which they are trained. Existing research indicates that humans may unconsciously internalize these biases, which can persist even after they stop using the programs. This study explores the cultural self-perception of LLMs by prompting ChatGPT (OpenAI) and Bard (Google) with value questions derived from the GLOBE project. The findings reveal that their cultural self-perception is most closely aligned with the values of English-speaking countries and countries characterized by sustained economic competitiveness. Recognizing the cultural biases of LLMs and understanding how they work is crucial for all members of society because one does not want the black box of artificial intelligence to perpetuate bias in humans, who might, in turn, inadvertently create and train even more biased algorithms.
20 pages, 3 tables, 4 figures; Online Supplement: 10 pages, 5 tables, 3 figures
References in corpus (8)
- Language Models are Few-Shot Learners
- Sparks of Artificial General Intelligence: Early experiments with GPT-4
- Aligning Large Language Models with Human: A Survey
- AI in the Gray: Exploring Moderation Policies in Dialogic Large Language Models vs. Human Answers in Controversial Topics
- Intersectional Bias in Causal Language Models
- Unveiling Gender Bias in Terms of Profession Across LLMs: Analyzing and Addressing Sociological Implications
- Cultural Alignment in Large Language Models: An Explanatory Analysis Based on Hofstede's Cultural Dimensions
- CBBQ: A Chinese Bias Benchmark Dataset Curated with Human-AI Collaboration for Large Language Models