2 papers
cs.CL2026
RMGAP: Benchmarking the Generalization of Reward Models across Diverse Preferences
Yangyang Zhou, Yi-Chen Li
Reinforcement Learning from Human Feedback has become the standard paradigm for language model alignment, where reward models directly determine alignment effectiveness. In this wo…
cs.HC2025
CPVis: Evidence-based Multimodal Learning Analytics for Evaluation in Collaborative Programming
Gefei Zhang, Shenming Ji, Yicao Li +5
As programming education becomes more widespread, many college students from non-computer science backgrounds begin learning programming. Collaborative programming emerges as an ef…