2026
Tab-MIA: A Benchmark Dataset for Membership Inference Attacks on Tabular Data in LLMs
ICLR 2026poster
Large language models (LLMs) are increasingly trained on tabular data, which, unlike unstructured text, often contains personally identifiable information (PII) in a highly structured and explicit format. As a result, privacy risks arise, since sensitive records can be inadvertently retained by the…