← Search

Hanyin Shao

2 accepted papers

2022

Are Large Pre-Trained Language Models Leaking Your Personal Information?

EMNLP 2022finding

Are Large Pre-Trained Language Models Leaking Your Personal Information? In this paper, we analyze whether Pre-Trained Language Models (PLMs) are prone to leaking personal information. Specifically, we query PLMs for email addresses with contexts of the email address or prompts containing the owner’…

2022

Understanding Jargon: Combining Extraction and Generation for Definition Modeling

EMNLP 2022main

Can machines know what twin prime is? From the composition of this phrase, machines may guess twin prime is a certain kind of prime, but it is still difficult to deduce exactly what twin stands for without additional knowledge. Here, twin prime is a jargon - a specialized term used by experts in a p…