1 paper · 1 filter
Yuanning Feng, Sinan Wang, Zhengxiang Cheng +2
LLM-as-a-Judge has been widely adopted as an evaluation method and served as supervised rewards in model training. However, existing benchmarks for LLM-as-a-Judge are mainly relyin…