← Search

KiYoon Yoo

5 accepted papers

2025

Nearly Zero-Cost Protection Against Mimicry by Personalized Diffusion Models

CVPR 2025poster

Recent advancements in diffusion models revolutionize image generation but pose risks of misuse, such as replicating artworks or generating deepfakes. Existing image protection methods, though effective, struggle to balance protection efficacy, invisibility, and latency, thus limiting practical use.…

Cited by 0SourcePDFScholar
2024

Advancing Beyond Identification: Multi-bit Watermark for Large Language Models

NAACL 2024long

We show the viability of tackling misuses of large language models beyond the identification of machine-generated text. While existing zero-bit watermark methods focus on detection only, some malicious misuses demand tracing the adversary user for counteracting them. To address this, we propose Mult…

2023

Robust Multi-bit Natural Language Watermarking through Invariant Features

ACL 2023long

Recent years have witnessed a proliferation of valuable original natural language contents found in subscription-based media outlets, web novel platforms, and outputs of large language models. However, these contents are susceptible to illegal piracy and potential misuse without proper security meas…

2022

Detection of Adversarial Examples in Text Classification: Benchmark and Baseline via Robust Density Estimation

ACL 2022findings

Word-level adversarial attacks have shown success in NLP models, drastically decreasing the performance of transformer-based models in recent years. As a countermeasure, adversarial defense has been explored, but relatively few efforts have been made to detect adversarial examples. However, detectin…