Showing cs.CLShow all
3 papers · 1 filter
cs.CL2026
LLMs as Span Annotators: A Comparative Study of LLMs and Humans
ZdenÄk Kasner, Vilém Zouhar, PatrÃcia Schmidtová +7
Span annotation - annotating specific text features at the span level - can be used to evaluate texts where single-score metrics fail to provide actionable feedback. Until recently…
cs.CL2024
Automatic Metrics in Natural Language Generation: A Survey of Current Evaluation Practices
PatrÃcia Schmidtová, Saad Mahamood, Simone Balloccu +6
Automatic metrics are extensively used to evaluate natural language processing systems. However, there has been increasing focus on how they are used and reported by practitioners…
cs.CL2024
factgenie: A Framework for Span-based Evaluation of Generated Texts
ZdenÄk Kasner, OndÅej Plátek, PatrÃcia Schmidtová +2
We present factgenie: a framework for annotating and visualizing word spans in textual model outputs. Annotations can capture various span-based phenomena such as semantic inaccura…