← Search

Jiyang Guan

4 accepted papers

2025

Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?

CVPR 2025poster

Multi-modal large language models (MLLMs) have made significant progress, yet their safety alignment remains limited. Typically, current open-source MLLMs rely on the alignment inherited from their language module to avoid harmful generations. However, the lack of safety measures specifically design…

Cited by 1SourcePDFScholar
2022

Are You Stealing My Model? Sample Correlation for Fingerprinting Deep Neural Networks

NeurIPS 2022accept

An off-the-shelf model as a commercial service could be stolen by model stealing attacks, posing great threats to the rights of the model owner. Model fingerprinting aims to verify whether a suspect model is stolen from the victim model, which gains more and more attention nowadays. Previous methods…