← Search

Yurui Chang

3 accepted papers

2025

AdvI2I: Adversarial Image Attack on Image-to-Image Diffusion Models

ICML 2025poster

Recent advances in diffusion models have significantly enhanced the quality of image synthesis, yet they have also introduced serious safety concerns, particularly the generation of Not Safe for Work (NSFW) content. Previous research has demonstrated that adversarial prompts can be used to generate…

2025

JoPA: Explaining Large Language Model’s Generation via Joint Prompt Attribution

ACL 2025long

Large Language Models (LLMs) have demonstrated impressive performances in complex text generation tasks. However, the contribution of the input prompt to the generated content still remains obscure to humans, underscoring the necessity of understanding the causality between input and output pairs. E…

2025

Monitoring Decoding: Mitigating Hallucination via Evaluating the Factuality of Partial Response during Generation

ACL 2025finding

While large language models have demonstrated exceptional performance across a wide range of tasks, they remain susceptible to hallucinations – generating plausible yet factually incorrect contents. Existing methods to mitigating such risk often rely on sampling multiple full-length generations, whi…

Cited by 0SourcePDFScholar