1 paper
Dani Roytburg, Matthew Bozoukov, Matthew Nguyen +3
Recent research has shown that large language models (LLMs) favor their own outputs when acting as judges, undermining the integrity of automated post-training and evaluation workf…