Uncovering Latent Memories: Assessing Data Leakage and Memorization Patterns in Large Language Models

Duan, Sunny; Khona, Mikail; Iyer, Abhiram; Schaeffer, Rylan; Fiete, Ila R

Computer Science > Computer Vision and Pattern Recognition

arXiv:2406.14549v1 (cs)

[Submitted on 20 Jun 2024 (this version), latest version 25 Jul 2024 (v2)]

Title:Uncovering Latent Memories: Assessing Data Leakage and Memorization Patterns in Large Language Models

Authors:Sunny Duan, Mikail Khona, Abhiram Iyer, Rylan Schaeffer, Ila R Fiete

View PDF HTML (experimental)

Abstract:The proliferation of large language models has revolutionized natural language processing tasks, yet it raises profound concerns regarding data privacy and security. Language models are trained on extensive corpora including potentially sensitive or proprietary information, and the risk of data leakage -- where the model response reveals pieces of such information -- remains inadequately understood. This study examines susceptibility to data leakage by quantifying the phenomenon of memorization in machine learning models, focusing on the evolution of memorization patterns over training. We investigate how the statistical characteristics of training data influence the memories encoded within the model by evaluating how repetition influences memorization. We reproduce findings that the probability of memorizing a sequence scales logarithmically with the number of times it is present in the data. Furthermore, we find that sequences which are not apparently memorized after the first encounter can be uncovered throughout the course of training even without subsequent encounters. The presence of these latent memorized sequences presents a challenge for data privacy since they may be hidden at the final checkpoint of the model. To this end, we develop a diagnostic test for uncovering these latent memorized sequences by considering their cross entropy loss.

Subjects:	Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Neurons and Cognition (q-bio.NC)
Cite as:	arXiv:2406.14549 [cs.CV]
	(or arXiv:2406.14549v1 [cs.CV] for this version)
	https://doi.org/10.48550/arXiv.2406.14549

Submission history

From: Sunny Duan [view email]
[v1] Thu, 20 Jun 2024 17:56:17 UTC (3,048 KB)
[v2] Thu, 25 Jul 2024 14:33:33 UTC (3,854 KB)

Computer Science > Computer Vision and Pattern Recognition

Title:Uncovering Latent Memories: Assessing Data Leakage and Memorization Patterns in Large Language Models

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Computer Vision and Pattern Recognition

Title:Uncovering Latent Memories: Assessing Data Leakage and Memorization Patterns in Large Language Models

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators