3 citations · 4 across the 2 of their papers we have counts for
1 paper · 1 filter
Andrea Sottana, Bin Liang, Kai Zou +1
Large Language Models (LLMs) evaluation is a patchy and inconsistent landscape, and it is becoming clear that the quality of automatic evaluation metrics is not keeping up with the…