1 paper · 1 filter
Himil Vasava, Ming Jiang
LLM-based evaluators of natural language generation (NLG) quality are widely deployed as scoring tools and as automated training signals, yet the internal procedure by which they a…