← Search

Yazhe Wang

2 accepted papers

2025

Automated Detection of Pre-training Text in Black-box LLMs

IJCAI 2025

Detecting whether a given text is a member in the pre-training data of Large Language Models (LLMs) is crucial for ensuring data privacy and copyright protection. Most existing methods rely on the LLM's hidden information (e.g., model parameters or token probabilities), making them ineffective in th

2024

NAPGuard: Towards Detecting Naturalistic Adversarial Patches

CVPR 2024poster

Recently the emergence of naturalistic adversarial patch (NAP) which possesses a deceptive appearance and various representations underscores the necessity of developing robust detection strategies. However existing approaches fail to differentiate the deep-seated natures in adversarial patches i.e.…

Cited by 8SourcePDFScholar