2 papers
cs.CL2026
ADE: Agentic Data Evolution Framework for Human-Centered Objectives
Yang Yu, Yilin Jiang, Zexuan Fei +6
Aligning large language models to human-centered objectives is difficult when targets are non-executable and context-dependent, limiting reliable verification and scalable supervis…
cs.CL2025
Teaching LLMs for Step-Level Automatic Math Correction via Reinforcement Learning
Junsong Li, Jie Zhou, Yutao Yang +7
Automatic math correction aims to check students' solutions to mathematical problems via artificial intelligence technologies. Most existing studies focus on judging the final answ…