2 papers
cs.CL2026
ELBench: A Multi-Dimensional Benchmark for Education-Facing Large Language Models
Yilin Jiang, Xiaorong Zhu, Fei Tan +9
Large language models are increasingly deployed in education as tutors, teaching assistants, and content generators. These roles place demands that ordinary question answering does…
cs.CY2026
Measuring the Professional Educational Competence of Foundation Models
Keqian Li, Mingzi Zhang, Xiaolong Wang +1
Foundation models tutor, assess, and instruct at population scale, requiring externally defined measures of educational competence. Existing benchmarks emphasize difficult academic…