← Search

Jiankun Peng

1 accepted papers

2025

Automated Detection of Pre-training Text in Black-box LLMs

IJCAI 2025

Detecting whether a given text is a member in the pre-training data of Large Language Models (LLMs) is crucial for ensuring data privacy and copyright protection. Most existing methods rely on the LLM's hidden information (e.g., model parameters or token probabilities), making them ineffective in th