Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Training on the Benchmark Is Not All You Need
Shiwen Ni, Xiangtao Kong, Chengming Li +4
The success of Large Language Models (LLMs) relies heavily on the huge amount of pre-training data learned in the pre-training phase. The opacity of the pre-training process and th…
cs.CL2024
E-EVAL: A Comprehensive Chinese K-12 Education Evaluation Benchmark for Large Language Models
Jinchang Hou, Chang Ao, Haihong Wu +8
With the accelerating development of Large Language Models (LLMs), many LLMs are beginning to be used in the Chinese K-12 education domain. The integration of LLMs and education is…