2 papers
cs.CL2025
Protecting De-identified Documents from Search-based Linkage Attacks
Pierre Lison, Mark Anderson
While de-identification models can help conceal the identity of the individuals mentioned in a document, they fail to address linkage risks, defined as the potential to map the de-…
cs.CL2024
Truthful Text Sanitization Guided by Inference Attacks
Ildikó Pilán, Benet Manzanares-Salor, David Sánchez +1
Text sanitization aims to rewrite parts of a document to prevent disclosure of personal information. The central challenge of text sanitization is to strike a balance between priva…