1 paper
Zailong Tian, Zhuoheng Han, Yanzhe Chen +5
Large Language Models (LLMs) are widely used as automated judges, where practical value depends on both accuracy and trustworthy, risk-aware judgments. Existing approaches predomin…