1 paper
Gerrit J. J. van den Burg, Gen Suzuki, Wei Liu +1
Large language models (LLMs) are increasingly used as automated judges to evaluate recommendation systems, search engines, and other subjective tasks, where relying on human evalua…