ICML 2026poster0 citations

Position: Generative Distributional Integrity against Backdoor Attacks

Shuaibiao Han, Ruiyang Ni, Zhiguo Yang, Changlong Li, Perley Xu, Wenjie Ruan

Abstract

Foundation models, such as Diffusion Models (DMs) and Large Language Models (LLMs), are now widely integrated into digital systems. This widespread use introduces a specific security risk: generative backdoors. Unlike traditional models where backdoors cause simple classification errors, generative backdoors hide within the model’s output distribution. This makes them difficult to detect using standard pattern-based methods.This paper argues that current defensive strategies are insufficient for generative AI. \textbf{We propose Distributional Integrity, a framework that focuses on maintaining the stability and accuracy of the model's data distribution.} We identify two primary threats: backdoors within the model supply chain and the contamination of synthetic data pipelines. To address these, we advocate for a shift toward cross-modal certification and parameter-level verification. These methods aim to secure the AI-generated content (AIGC) ecosystem against inherited vulnerabilities.

LLMDiffusion
BibTeX
@inproceedings{icml2026_positiongenerati,
  title = {Position: Generative Distributional Integrity against Backdoor Attacks},
  author = {Shuaibiao Han and Ruiyang Ni and Zhiguo Yang and Changlong Li and Perley Xu and Wenjie Ruan},
  booktitle = {ICML 2026},
  year = {2026}
}