DF-MIA: A Distribution-Free Membership Inference Attack on Fine-Tuned Large Language Models
Membership Inference Attack (MIA) aims to determine if a specific sample is present in the training dataset of a target machine learning model. Previous MIAs against fine-tuned Large Language Models (LLMs) either fail to address the unique challenges in the fine-tuned setting or rely on strong assu…