← Search

Tianshuo Cong

4 accepted papers

2025

Beyond the Tip of Efficiency: Uncovering the Submerged Threats of Jailbreak Attacks in Small Language Models

ACL 2025finding

Small language models (SLMs) have become increasingly prominent in the deployment on edge devices due to their high efficiency and low computational cost. While researchers continue to advance the capabilities of SLMs through innovative training strategies and model compression techniques, the secur…

Cited by 0SourcePDFScholar
2025

CL-Attack: Textual Backdoor Attacks via Cross-Lingual Triggers

AAAI 2025technical

Backdoor attacks significantly compromise the security of large language models by triggering them to output specific and controlled content. Currently, triggers for textual backdoor attacks fall into two categories: fixed-token triggers and sentence-pattern triggers. However, the former are typical…

2025

ErrorTrace: A Black-Box Traceability Mechanism Based on Model Family Error Space

NeurIPS 2025spotlight

The open-source release of large language models (LLMs) enables malicious users to create unauthorized derivative models at low cost, posing significant threats to intellectual property (IP) and market stability. Existing IP protection methods either require access to model parameters or are vulnera…

Cited by 0SourcecodeScholar
2025

FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts

AAAI 2025technical

Large Vision-Language Models (LVLMs) signify a groundbreaking paradigm shift within the Artificial Intelligence (AI) community, extending beyond the capabilities of Large Language Models (LLMs) by assimilating additional modalities (e.g., images). Despite this advancement, the safety of LVLMs remain…