Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Ruzzetti, Elena Sofia; Xompero, Giancarlo A.; Venditti, Davide; Zanzotto, Fabio Massimo

doi:10.18653/v1/2025.acl-long.810

Computer Science > Cryptography and Security

arXiv:2506.10024 (cs)

[Submitted on 9 Jun 2025]

Title:Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Authors:Elena Sofia Ruzzetti, Giancarlo A. Xompero, Davide Venditti, Fabio Massimo Zanzotto

View PDF HTML (experimental)

Abstract:Large Language Models (LLMs) memorize, and thus, among huge amounts of uncontrolled data, may memorize Personally Identifiable Information (PII), which should not be stored and, consequently, not leaked. In this paper, we introduce Private Memorization Editing (PME), an approach for preventing private data leakage that turns an apparent limitation, that is, the LLMs' memorization ability, into a powerful privacy defense strategy. While attacks against LLMs have been performed exploiting previous knowledge regarding their training data, our approach aims to exploit the same kind of knowledge in order to make a model more robust. We detect a memorized PII and then mitigate the memorization of PII by editing a model knowledge of its training data. We verify that our procedure does not affect the underlying language model while making it more robust against privacy Training Data Extraction attacks. We demonstrate that PME can effectively reduce the number of leaked PII in a number of configurations, in some cases even reducing the accuracy of the privacy attacks to zero.

Comments:	To be published at ACL 2025 (Main)
Subjects:	Cryptography and Security (cs.CR); Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
Cite as:	arXiv:2506.10024 [cs.CR]
	(or arXiv:2506.10024v1 [cs.CR] for this version)
	https://doi.org/10.48550/arXiv.2506.10024
Related DOI:	https://doi.org/10.18653/v1/2025.acl-long.810

Submission history

From: Elena Sofia Ruzzetti [view email]
[v1] Mon, 9 Jun 2025 17:57:43 UTC (8,890 KB)

Computer Science > Cryptography and Security

Title:Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Cryptography and Security

Title:Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators