Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
OmniEduBench: A Comprehensive Chinese Benchmark for Evaluating Large Language Models in Education
Min Zhang, Hao Chen, Wenqi Zhang +6
With the rapid development of large language models (LLMs), various LLM-based works have been widely applied in educational fields. However, most existing LLMs and their benchmarks…
cs.CL2024
DOP: Diagnostic-Oriented Prompting for Large Language Models in Mathematical Correction
Hao Chen, Biaojie Zeng, Xin Lin +2
Math world problems correction(MWPC) is a novel task dedicated to rectifying reasoning errors in the process of solving mathematical problems. In this paper, leveraging the advance…