Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Position: AI Evaluation Should Learn from How We Test Humans
Yan Zhuang, Qi Liu, Zachary A. Pardos +5
As AI systems continue to evolve, their rigorous evaluation becomes crucial for their development and deployment. Researchers have constructed various large-scale benchmarks to det…
cs.CL2024
EduNLP: Towards a Unified and Modularized Library for Educational Resources
Zhenya Huang, Yuting Ning, Longhu Qin +8
Educational resource understanding is vital to online learning platforms, which have demonstrated growing applications recently. However, researchers and developers always struggle…