2 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Hongli Li, Che Han Chen, Kevin Fan +3
Despite the growing promise of large language models (LLMs) in automated essay scoring (AES), empirical findings regarding their reliability compared to human raters remain mixed.…