1 paper · 1 filter
Kimberly Le Truong, Riccardo Fogliato, Hoda Heidari +1
Current benchmarks for evaluating Large Language Models (LLMs) often do not exhibit enough writing style diversity, with many adhering primarily to standardized conventions. Such b…