← Search

Pei Yang

8 accepted papers

2025

IDProtector: An Adversarial Noise Encoder to Protect Against ID-Preserving Image Generation

CVPR 2025poster

Recently, zero-shot methods like InstantID have revolutionized identity-preserving generation. Unlike multi-image finetuning approaches such as DreamBooth, these zero-shot methods leverage powerful facial encoders to extract identity information from a single portrait photo, enabling efficient ident…

2025

Multi-label body constitution recognition via dual transform MLP-like architecture using tongue images

ICASSP 2025accepted

According to the traditional Chinese medicine composite constitution theory, body constitution recognition is modeled as a unique task of multi-label problems using tongue images. Although MLP-like architecture is one of the base models except for CNN-based, Transformer-based, and MLP-like-based for…

Cited by 0SourceScholar
2025

WMAdapter: Adding WaterMark Control to Latent Diffusion Models

ICML 2025poster

Watermarking is essential for protecting the copyright of AI-generated images. We propose WMAdapter, a diffusion model watermark plugin that embeds user-specified watermark information seamlessly during the diffusion generation process. Unlike previous methods that modify diffusion modules to incorp…

Cited by 14SourcePDFScholar
2024

Rethinking Tree-Ring Watermarking for Enhanced Multi-Key Identification

ECCV 2024poster

"We revisit Tree-Ring Watermarking, a recent diffusion model watermarking method that demonstrates great robustness to various attacks. We conduct an in-depth study on it and reveal that the distribution shift unintentionally introduced by the watermarking process, apart from watermark pattern match…

2024

Speech Relationship Learning for Cross-Corpus Speech Emotion Recognition

ICASSP 2024accepted

Cross-Corpus Speech Emotion Recognition (SER) aims to identify human emotions from speech across different speakers and languages. Previous work engaged in extracting the domain-invariant features among individual samples that are most relevant to emotions, ignoring rich relationships between speech…

Cited by 0SourceScholar