1 paper
Gavin Abercrombie, Tanvi Dinkar, Amanda Cercas Curry +2
We commonly use agreement measures to assess the utility of judgements made by human annotators in Natural Language Processing (NLP) tasks. While inter-annotator agreement is frequ…