2 papers
cs.CL2023
Neural Text Sanitization with Privacy Risk Indicators: An Empirical Analysis
Anthi Papadopoulou, Pierre Lison, Mark Anderson +2
Text sanitization is the task of redacting a document to mask all occurrences of (direct or indirect) personal identifiers, with the goal of concealing the identity of the individu…
cs.CL2023
Conversational Feedback in Scripted versus Spontaneous Dialogues: A Comparative Analysis
Ildikó Pilán, Laurent Prévot, Hendrik Buschmeier +1
Scripted dialogues such as movie and TV subtitles constitute a widespread source of training data for conversational NLP models. However, there are notable linguistic differences b…