17 citations · 33 across the 2 of their papers we have counts for
1 paper · 1 filter
Elizabeth Clark, Tal August, Sofia Serrano +3
Human evaluations are typically considered the gold standard in natural language generation, but as models' fluency improves, how well can evaluators detect and judge machine-gener…