← Search

Datao You

2 accepted papers

2025

Exploring Jailbreak Attacks on LLMs through Intent Concealment and Diversion

ACL 2025finding

Although large language models (LLMs) have achieved remarkable advancements, their security remains a pressing concern. One major threat is jailbreak attacks, where adversarial prompts bypass model safeguards to generate harmful or objectionable content. Researchers study jailbreak attacks to unders…

Cited by 0SourcePDFScholar
2025

Low-Resource Fast Text Classification Based on Intra-Class and Inter-Class Distance Calculation

COLING 2025main

In recent years, text classification methods based on neural networks and pre-trained models have gained increasing attention and demonstrated excellent performance. However, these methods still have some limitations in practical applications: (1) They typically focus only on the matching similarity…

Cited by 1SourcePDFScholar