From the 1 of 4 linked papers with an AI index.
4 papers
WrAFT: a Modularized Automated Writing Evaluation System for Argumentative Essays
Adnan Labib, Yixuan Huang, Jiahui Wu +3
The paper presents WrAFT, a modular system for automated evaluation of argumentative essays that provides scoring and both surface‑level and deep‑level feedback using large languag…
Multi-Dimensional Evaluation of LLMs for Grammatical Error Correction
Adnan Labib, Qiao Wang, Yixuan Huang +1
Automated assistants for Grammatical Error Correction are now embedded in educational platforms serving millions of learners, yet three critical gaps remain in this domain: (1) lat…
GenQuest: An LLM-based Text Adventure Game for Language Learners
Qiao Wang, Adnan Labib, Robert Swier +2
GenQuest is a generative text adventure game that leverages Large Language Models (LLMs) to facilitate second language learning through immersive, interactive storytelling. The sys…
Beyond the Score: Uncertainty-Calibrated LLMs for Automated Essay Assessment
Ahmed Karim, Qiao Wang, Zheng Yuan
Automated Essay Scoring (AES) systems now reach near human agreement on some public benchmarks, yet real-world adoption, especially in high-stakes examinations, remains limited. A…