4 papers
Unveiling Scoring Processes: Dissecting the Differences between LLMs and Human Graders in Automatic Scoring
Xuansheng Wu, Padmaja Pravin Saraf, Gyeonggeon Lee +3
Large language models (LLMs) have demonstrated strong potential in performing automatic scoring for constructed response assessments. While constructed responses graded by humans a…
Using GPT-4 to Augment Unbalanced Data for Automatic Scoring
Luyang Fang, Gyeong-Geon Lee, Xiaoming Zhai
Machine learning-based automatic scoring faces challenges with unbalanced student responses across scoring categories. To address this, we introduce a novel text data augmentation…
G-SciEdBERT: A Contextualized LLM for Science Assessment Tasks in German
Ehsan Latif, Gyeong-Geon Lee, Knut Neumann +2
The advancement of natural language processing has paved the way for automated scoring systems in various languages, such as German (e.g., German BERT [G-BERT]). Automatically scor…
Realizing Visual Question Answering for Education: GPT-4V as a Multimodal AI
Gyeong-Geon Lee, Xiaoming Zhai
Educational scholars have analyzed various image data acquired from teaching and learning situations, such as photos that shows classroom dynamics, students' drawings with regard t…