1 paper
Khaoula Chehbouni, Melina Medjdoub, Florian Carichon +2
In recent years, large language models (LLMs) have emerged as a popular alternative for evaluation. Often referred to as LLMs as judges (LLJs), these systems have been widely adopt…