1 paper · 1 filter
Wenbo Zhang, Lijinghua Zhang, Liner Xiang +1
Reasoning-capable large language models (LLMs) have recently been adopted as automated judges, but their benefits and costs in LLM-as-a-Judge settings remain unclear. Through contr…