1 paper · 1 filter
Guijin Son, Hyunwoo Ko, Hoyoung Lee +2
LLM-as-a-Judge and reward models are widely used alternatives of multiple-choice questions or human annotators for large language model (LLM) evaluation. Their efficacy shines in e…