← Search

Jiawei Ye

4 accepted papers

2026

StyleBreak: Revealing Alignment Vulnerabilities in Large Audio-Language Models via Style-Aware Audio Jailbreak

AAAI 2026technical

Large Audio-language Models (LAMs) have recently enabled powerful speech-based interactions by coupling audio encoders with Large Language Models (LLMs). However, the security of LAMs under adversarial attacks remains underexplored, especially through audio jailbreaks that craft malicious audio prom

Cited by 0SourcePDFScholar
2025

Efficient and Expandable Token-Level Approach for Multi-Domain Sensitive Information Classification

ICASSP 2025accepted

Incorporating privacy regulations and business requirements, enterprises should securely manage unstructured textual data from diverse domains. Sensitive information classification is a critical component of data security, but it poses challenges due to complex textual contexts. With the ever-increa…

Cited by 0SourceScholar
2025

JailPO: A Novel Black-Box Jailbreak Framework via Preference Optimization Against Aligned LLMs

AAAI 2025technical

Large Language Models (LLMs) aligned with human feedback have recently garnered significant attention. However, it remains vulnerable to jailbreak attacks, where adversaries manipulate prompts to induce harmful outputs. Exploring jailbreak attacks enables us to investigate the vulnerabilities of LLM…

Cited by 0SourcePDFScholar