2 papers
cs.CL2024
Standardizing the Measurement of Text Diversity: A Tool and a Comparative Analysis of Scores
Chantal Shaib, Venkata S. Govindarajan, Joe Barrow +4
The diversity across outputs generated by LLMs shapes perception of their quality and utility. High lexical diversity is often desirable, but there is no standard method to measure…
cs.CL2024
How Much Annotation is Needed to Compare Summarization Models?
Chantal Shaib, Joe Barrow, Alexa F. Siu +2
Modern instruction-tuned models have become highly capable in text generation tasks such as summarization, and are expected to be released at a steady pace. In practice one may now…