Practical and Ethical Challenges of Large Language Models in Education: A Systematic Scoping Review
arXiv:2303.13379 · doi:10.1111/bjet.13370
Abstract
Educational technology innovations leveraging large language models (LLMs) have shown the potential to automate the laborious process of generating and analysing textual content. While various innovations have been developed to automate a range of educational tasks (e.g., question generation, feedback provision, and essay grading), there are concerns regarding the practicality and ethicality of these innovations. Such concerns may hinder future research and the adoption of LLMs-based innovations in authentic educational contexts. To address this, we conducted a systematic scoping review of 118 peer-reviewed papers published since 2017 to pinpoint the current state of research on using LLMs to automate and support educational tasks. The findings revealed 53 use cases for LLMs in automating education tasks, categorised into nine main categories: profiling/labelling, detection, grading, teaching support, prediction, knowledge representation, feedback, content generation, and recommendation. Additionally, we also identified several practical and ethical challenges, including low technological readiness, lack of replicability and transparency, and insufficient privacy and beneficence considerations. The findings were summarised into three recommendations for future studies, including updating existing innovations with state-of-the-art models (e.g., GPT-3/4), embracing the initiative of open-sourcing models/systems, and adopting a human-centred approach throughout the developmental process. As the intersection of AI and education is continuously evolving, the findings of this study can serve as an essential reference point for researchers, allowing them to leverage the strengths, learn from the limitations, and uncover potential research opportunities enabled by ChatGPT and other generative AI models.
References in corpus (3)
Cited by in corpus (13)
- Generative Artificial Intelligence in Learning Analytics: Contextualising Opportunities and Challenges through the Learning Analytics Cycle
- Knowledge Graphs as Context Sources for LLM-Based Explanations of Learning Recommendations
- Could ChatGPT get an Engineering Degree? Evaluating Higher Education Vulnerability to AI Assistants
- Chatbot-supported Thesis Writing: An Autoethnographic Report
- Towards responsible AI for education: Hybrid human-AI to confront the Elephant in the room
- Generative AI and Agency in Education: A Critical Scoping Review and Thematic Analysis
- Annotation Guidelines-Based Knowledge Augmentation: Towards Enhancing Large Language Models for Educational Text Classification
- Agentic AI in Healthcare & Medicine: A Seven-Dimensional Taxonomy for Empirical Evaluation of LLM-based Agents
- Can GPT-4o Evaluate Usability Like Human Experts? A Comparative Study on Issue Identification in Heuristic Evaluation
- Evaluating Interactivity: Toward Automated Assessment of AI-Generated Explorable Explanations
- Ethical AI prompt recommendations in large language models using collaborative filtering
- Methodologies for Improving the Quality of AI Tutoring in K-12 Education
- AI of the People, by the People, for the People: A Social Choice Approach to Collective Control of Artificial Intelligence