6 citations · 6 across the 1 of their papers we have counts for
2 papers
cs.AI2026
LLMs Do Not Grade Essays Like Humans
Jerin George Mathew, Sumayya Taher, Anindita Kundu +1
Large language models have recently been proposed as tools for automated essay scoring, but their agreement with human grading remains unclear. In this work, we evaluate how LLM-ge…
cs.CL2024★ 6 cited
Are Large Language Models Good Essay Graders?
Anindita Kundu, Denilson Barbosa
We evaluate the effectiveness of Large Language Models (LLMs) in assessing essay quality, focusing on their alignment with human grading. More precisely, we evaluate ChatGPT and Ll…