1 paper
Yasi Zhang, Tianyu Chen, Mingyuan Zhou +3
Large language models (LLMs) are increasingly deployed as automated evaluators that assign numeric scores to model outputs, a paradigm known as LLM-as-a-Judge. However, standard Re…