On Protecting the Data Privacy of Large Language Models (LLMs): A Survey

Biwei Yan; Kun Li; Minghui Xu; Yueyan Dong; Yue Zhang; Zhaochun Ren; Xiuzhen Cheng

doi:10.48550/arxiv.2403.05156

Back

On Protecting the Data Privacy of Large Language Models (LLMs): A Survey

Preprint

Open access

On Protecting the Data Privacy of Large Language Models (LLMs): A Survey

Biwei Yan, Kun Li, Minghui Xu, Yueyan Dong, Yue Zhang, Zhaochun Ren and Xiuzhen Cheng

arXiv.org

14 Mar 2024

DOI: https://doi.org/10.48550/arxiv.2403.05156

Files and links (1)

url

https://doi.org/10.48550/arxiv.2403.05156View

Preprint (Author's original)arXiv.org - Non-exclusive license to distribute, Open

Abstract

Computer Science - Cryptography and Security

Large language models (LLMs) are complex artificial intelligence systems capable of understanding, generating and translating human language. They learn language patterns by analyzing large amounts of text data, allowing them to perform writing, conversation, summarizing and other language tasks. When LLMs process and generate large amounts of data, there is a risk of leaking sensitive information, which may threaten data privacy. This paper concentrates on elucidating the data privacy concerns associated with LLMs to foster a comprehensive understanding. Specifically, a thorough investigation is undertaken to delineate the spectrum of data privacy threats, encompassing both passive privacy leakage and active privacy attacks within LLMs. Subsequently, we conduct an assessment of the privacy protection mechanisms employed by LLMs at various stages, followed by a detailed examination of their efficacy and constraints. Finally, the discourse extends to delineate the challenges encountered and outline prospective directions for advancement in the realm of LLM privacy protection.

Metrics

32 Record Views

Details

Title: On Protecting the Data Privacy of Large Language Models (LLMs): A Survey
Creators: Biwei Yan
Kun Li
Minghui Xu
Yueyan Dong
Yue Zhang
Zhaochun Ren
Xiuzhen Cheng
Publication Details: arXiv.org
Resource Type: Preprint
Language: English
Academic Unit: Computer Science (Computing)
Other Identifier: 991021871463204721

On Protecting the Data Privacy of Large Language Models (LLMs): A Survey

Files and links (1)

Abstract

Metrics

Details

Drexel University Social media