Return
Detecting Training Data For Large Language Models: A Survey
C
J
S
Y
H
X
D
L
DOI:10.1145/3779430.png)
Abstract
En 中文
As large language models (LLMs) continue to evolve, the scope and diversity of data used for training are expanding significantly. However, the training dataset of LLMs may inevitably contain sensitive information such as personal data or copyrighted material, leading to privacy leakage or copyright infringement risks if the model generates highly similar or identical text to these sources. This has drawn attention to the issue of detecting whether the text data is used for LLM training. To date, research on detecting training data usage in artificial intelligence (AI) models has mainly focused on traditional machine learning (ML) models. However, studies on LLMs remain relatively immature. The lack of understanding of research progress in this area has hindered the development of more effective detection methods. Therefore, this article aims to address this gap by conducting the analysis of detecting training data for LLM. Specifically, we analyze the available LLM's information to the detector, the main detection methods, determination metrics, and discuss the technical challenges and potential directions for future research in this field.
Keywords:
Large language models
detecting training data
Journal
IF:
28
Papers:
2.4K
Citations:
3.5W
