1 paper · 1 filter
Melanie Subbiah, Faisal Ladhak, Akankshya Mishra +3
Human evaluation has been the gold standard for checking faithfulness in abstractive summarization. However, with a challenging source domain like narrative, multiple annotators ca…