Showing cs.CLShow all
3 papers · 1 filter
cs.CL2025
SMRC: Aligning Large Language Models with Student Reasoning for Mathematical Error Correction
Biaojie Zeng, Min Zhang, Juan Zhou +3
Large language models (LLMs) often make reasoning errors when solving mathematical problems, and how to automatically detect and correct these errors has become an important resear…
cs.CL2025
OmniEduBench: A Comprehensive Chinese Benchmark for Evaluating Large Language Models in Education
Min Zhang, Hao Chen, Wenqi Zhang +6
With the rapid development of large language models (LLMs), various LLM-based works have been widely applied in educational fields. However, most existing LLMs and their benchmarks…
cs.CL2025
EduDial: Constructing a Large-scale Multi-turn Teacher-Student Dialogue Corpus
Shouang Wei, Min Zhang, Xin Lin +3
Recently, several multi-turn dialogue benchmarks have been proposed to evaluate the conversational abilities of large language models (LLMs). As LLMs are increasingly recognized as…