← Search

Shuo Shao

3 accepted papers

2026

JANUS: A Lightweight Framework for Jailbreaking Text-to-Image Models via Distribution Optimization

CVPR 2026

Text-to-image (T2I) models such as Stable Diffusion and DALLE remain susceptible to generating harmful or Not-Safe-For-Work (NSFW) content under jailbreak attacks despite deployed safety filters. Existing jailbreak attacks either rely on proxy-loss optimization instead of the true end-to-end objecti

Cited by 0SourcecodeScholar
2026

MAJIC: Markovian Adaptive Jailbreaking via Iterative Composition of Diverse Innovative Strategies

AAAI 2026technical

Large Language Models (LLMs) have exhibited remarkable capabilities but remain vulnerable to jailbreaking attacks, which can elicit harmful content from the models by manipulating the input prompts. Existing black-box jailbreaking techniques primarily rely on static prompts crafted with a single, no

Cited by 0SourcePDFScholar
2025

REFINE: Inversion-Free Backdoor Defense via Model Reprogramming

ICLR 2025poster

Backdoor attacks on deep neural networks (DNNs) have emerged as a significant security threat, allowing adversaries to implant hidden malicious behaviors during the model training phase. Pre-processing-based defense, which is one of the most important defense paradigms, typically focuses on input tr…

Cited by 2SourcePDFScholar