Should ChatGPT be Biased? Challenges and Risks of Bias in Large Language Models
arXiv:2304.03738 · doi:10.5210/fm.v28i11.13346
Abstract
As the capabilities of generative language models continue to advance, the implications of biases ingrained within these models have garnered increasing attention from researchers, practitioners, and the broader public. This article investigates the challenges and risks associated with biases in large-scale language models like ChatGPT. We discuss the origins of biases, stemming from, among others, the nature of training data, model specifications, algorithmic constraints, product design, and policy decisions. We explore the ethical concerns arising from the unintended consequences of biased model outputs. We further analyze the potential opportunities to mitigate biases, the inevitability of some biases, and the implications of deploying these models in various applications, such as virtual assistants, content generation, and chatbots. Finally, we review the current approaches to identify, quantify, and mitigate biases in language models, emphasizing the need for a multi-disciplinary, collaborative effort to develop more equitable, transparent, and responsible AI systems. This article aims to stimulate a thoughtful dialogue within the artificial intelligence community, encouraging researchers and developers to reflect on the role of biases in generative language models and the ongoing pursuit of ethical AI.
Published on First Monday https://firstmonday.org/ojs/index.php/fm/article/view/13346/11365
References in corpus (4)
Cited by in corpus (16)
- Generative AI
- Fairness And Bias in Artificial Intelligence: A Brief Survey of Sources, Impacts, And Mitigation Strategies
- ChatGPT: Jack of all trades, master of none
- Cultural Bias and Cultural Alignment of Large Language Models
- GenAI Against Humanity: Nefarious Applications of Generative Artificial Intelligence and Large Language Models
- Factuality Challenges in the Era of Large Language Models
- A Comprehensive Survey of Large Language Models and Multimodal Large Language Models in Medicine
- Large Language Models for Wearable Sensor-Based Human Activity Recognition, Health Monitoring, and Behavioral Modeling: A Survey of Early Trends, Datasets, and Challenges
- An Autoethnographic Case Study of Generative Artificial Intelligence's Utility for Accessibility
- Emergent social conventions and collective bias in LLM populations
- Can GPT-4 learn to analyse moves in research article abstracts?
- "She was useful, but a bit too optimistic": Augmenting Design with Interactive Virtual Personas
- Information Suppression in Large Language Models: Auditing, Quantifying, and Characterizing Censorship in DeepSeek
- Benchmarking Adversarial Robustness to Bias Elicitation in Large Language Models: Scalable Automated Assessment with LLM-as-a-Judge
- The Strong Pull of Prior Knowledge in Large Language Models and Its Impact on Emotion Recognition
- Are Large Language Models Really Bias-Free? Jailbreak Prompts for Assessing Adversarial Robustness to Bias Elicitation