2025
Can We Infer Confidential Properties of Training Data from LLMs?
NeurIPS 2025spotlight
Large language models (LLMs) are increasingly fine-tuned on domain-specific datasets to support applications in fields such as healthcare, finance, and law. These fine-tuning datasets often have sensitive and confidential dataset-level properties — such as patient demographics or disease prevalence—…