Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Elena Sofia Ruzzetti; Giancarlo A. Xompero; Davide Venditti; Fabio Massimo Zanzotto

arXiv:2506.10024·cs.CR·August 22, 2025

Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models

Elena Sofia Ruzzetti, Giancarlo A. Xompero, Davide Venditti, Fabio Massimo Zanzotto

PDF

1 Repo 1 Video

TL;DR

This paper introduces Private Memorization Editing (PME), a novel method that leverages the memorization ability of large language models to prevent leakage of personally identifiable information, thereby enhancing data privacy.

Contribution

The paper proposes PME, a new technique that detects and edits memorized private data in LLMs without affecting their overall performance, strengthening privacy defenses.

Findings

01

PME effectively reduces PII leakage in various configurations.

02

In some cases, PME completely prevents privacy attacks.

03

PME maintains the underlying language model's performance.

Abstract

Large Language Models (LLMs) memorize, and thus, among huge amounts of uncontrolled data, may memorize Personally Identifiable Information (PII), which should not be stored and, consequently, not leaked. In this paper, we introduce Private Memorization Editing (PME), an approach for preventing private data leakage that turns an apparent limitation, that is, the LLMs' memorization ability, into a powerful privacy defense strategy. While attacks against LLMs have been performed exploiting previous knowledge regarding their training data, our approach aims to exploit the same kind of knowledge in order to make a model more robust. We detect a memorized PII and then mitigate the memorization of PII by editing a model knowledge of its training data. We verify that our procedure does not affect the underlying language model while making it more robust against privacy Training Data Extraction…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Code & Models

Repositories

elenasofia98/pme
pytorchOfficial

Videos

Private Memorization Editing: Turning Memorization into a Defense to Strengthen Data Privacy in Large Language Models· underline