1 paper · 1 filter
Gerrit J. J. van den Burg, Gen Suzuki, Wei Liu +1
Large language models (LLMs) are increasingly used as automated judges to evaluate recommendation systems, search engines, and other subjective tasks, where relying on human evalua…