← Search

Youngjin Jin

4 accepted papers

2025

Claim-Guided Textual Backdoor Attack for Practical Applications

NAACL 2025findings

Recent advances in natural language processing and the increased use of large language models have exposed new security vulnerabilities, such as backdoor attacks. Previous backdoor attacks require input manipulation after model distribution to activate the backdoor, posing limitations in real-world…

2024

Ignore Me But Don’t Replace Me: Utilizing Non-Linguistic Elements for Pretraining on the Cybersecurity Domain

NAACL 2024findings

Cybersecurity information is often technically complex and relayed through unstructured text, making automation of cyber threat intelligence highly challenging. For such text domains that involve high levels of expertise, pretraining on in-domain corpora has been a popular method for language models…

Cited by 3SourcePDFScholar
2023

DarkBERT: A Language Model for the Dark Side of the Internet

ACL 2023long

Recent research has suggested that there are clear differences in the language used in the Dark Web compared to that of the Surface Web. As studies on the Dark Web commonly require textual analysis of the domain, language models specific to the Dark Web may provide valuable insights to researchers.…

2022

Shedding New Light on the Language of the Dark Web

NAACL 2022long

The hidden nature and the limited accessibility of the Dark Web, combined with the lack of public datasets in this domain, make it difficult to study its inherent characteristics such as linguistic properties. Previous works on text classification of Dark Web domain have suggested that the use of de…

Cited by 17SourcePDFScholar