1 citations · 1 across the 1 of their papers we have counts for
1 paper
Andrea Sottana, Bin Liang, Kai Zou +1
Large Language Models (LLMs) evaluation is a patchy and inconsistent landscape, and it is becoming clear that the quality of automatic evaluation metrics is not keeping up with the…