Truthful text sanitization guided by inference attacks

Pilán, I; Manzanares-Salor, B; Sánchez, D; Lison, P

doi:10.1016/j.asoc.2025.114013

Identification data

Identifier: imarina:9467100

Handle: https://hdl.handle.net/20.500.11797/imarina9467100

Authors: Pilán, I; Manzanares-Salor, B; Sánchez, D; Lison, P

Abstract:
Text sanitization aims to rewrite parts of a document to prevent disclosure of personal information. The central challenge of text sanitization is to strike a balance between privacy protection (avoiding the leakage of personal information) and utility preservation (retaining as much as possible of the document's original content). To this end, we introduce a novel text sanitization method based on generalizations, that is, broader but still informative terms that subsume the semantic content of the original text spans. The approach relies on the use of instruction-tuned large language models (LLMs) and is divided into two stages. Given a document including text spans expressing personally identifiable information (PII), the LLM is first applied to obtain truth-preserving replacement candidates for each text span and rank them according to their abstraction level. Those candidates are then evaluated for their ability to protect privacy by conducting inference attacks with the LLM. Finally, the system selects the most informative replacement candidate shown to be resistant to those attacks. This two-stage process produces replacements that effectively balance privacy and utility. We also present novel metrics to evaluate these two aspects without needing to manually annotate documents. Results on the Text Anonymization Benchmark show that the proposed approach, implemented with Mistral 7B Instruct, leads to enhanced utility, with only a marginal ( < 1 p.p.) increase in re-identification risk compared to fully suppressing the original spans. Furthermore, our approach is shown to be more truth-preserving than existing methods such as Microsoft Presidio's synthetic replacements.
Others:

Link to the original source: https://www.sciencedirect.com/science/article/pii/S1568494625013262?via%3Dihub
APA: Pilán, I; Manzanares-Salor, B; Sánchez, D; Lison, P (2025). Truthful text sanitization guided by inference attacks. Applied Soft Computing, 185(), 114013-. DOI: 10.1016/j.asoc.2025.114013
Paper original source: Applied Soft Computing. 185 114013-
Article's DOI: 10.1016/j.asoc.2025.114013
Journal publication year: 2025-12-01
Entity: Universitat Rovira i Virgili
Paper version: info:eu-repo/semantics/publishedVersion
Record's date: 2026-02-13
URV's Author/s: Sánchez Ruenes, David
Department: Enginyeria Informàtica i Matemàtiques
Licence document URL: https://repositori.urv.cat/ca/proteccio-de-dades/
Publication Type: Journal Publications
Author, as appears in the article.: Pilán, I; Manzanares-Salor, B; Sánchez, D; Lison, P
licence for use: https://creativecommons.org/licenses/by/3.0/es/
Thematic Areas: Administração pública e de empresas, ciências contábeis e turismo, Biotecnología, Ciência da computação, Ciência de alimentos, Computer science, artificial intelligence, Computer science, interdisciplinary applications, Engenharias i, Engenharias ii, Engenharias iii, Engenharias iv, Interdisciplinar, Matemática / probabilidade e estatística, Software
Author's mail: david.sanchez@urv.cat

Keywords:

Agreement
Data privacy
Data utility
Information-content
Large language models
Redaction
Semantic similarity
Text sanitization
Truth-preserving replacements
Computer Science
Artificial Intelligence
Interdisciplinary Applications
Software
Administração pública e de empresas
ciências contábeis e turismo
Biotecnología
Ciência da computação
Ciência de alimentos
Engenharias i
Engenharias ii
Engenharias iii
Engenharias iv
Interdisciplinar
Matemática / probabilidade e estatística
Documents:

DocumentPrincipal
Cerca a google

Truthful text sanitization guided by inference attacks

Identification data

Others:

Keywords:

Documents:

Cerca a google