← Search

Dapeng Man

5 accepted papers

2025

Collapsing Sequence-Level Data-Policy Coverage via Poisoning Attack in Offline Reinforcement Learning

UAI 2025

Offline reinforcement learning (RL) heavily relies on the coverage of pre-collected data over the target policy’s distribution. Existing studies aim to improve data-policy coverage to mitigate distributional shifts, but overlook security risks from insufficient coverage, and the single-step analysis

Cited by 0SourcePDFScholar
2025

Robust Adversarial Training for Industrial Defect Classification with Long-Tailed Data

ICASSP 2025accepted

Deep neural networks are vulnerable to adversarial examples which fool model predictions by adding imperceptible perturbations to natural examples. Adversarial training is effective in defending against adversarial attacks but faces a challenge with long-tailed data, where the over-compression of ta…

Cited by 0SourceScholar
2024

Beyond Traditional Threats: A Persistent Backdoor Attack on Federated Learning

AAAI 2024technical

Backdoors on federated learning will be diluted by subsequent benign updates. This is reflected in the significant reduction of attack success rate as iterations increase, ultimately failing. We use a new metric to quantify the degree of this weakened backdoor effect, called attack persistence. Give…

2024

Bridging the Gaps of Both Modality and Language: Synchronous Bilingual CTC for Speech Translation and Speech Recognition

ICASSP 2024accepted

In this study, we present synchronous bilingual Connectionist Temporal Classification (CTC), an innovative framework that leverages dual CTC to bridge the gaps of both modality and language in the speech translation (ST) task. Utilizing transcript and translation as concurrent objectives for CTC, ou…

Cited by 0SourceScholar
2024

Revisiting Interpolation Augmentation for Speech-to-Text Generation

ACL 2024findings

Speech-to-text (S2T) generation systems frequently face challenges in low-resource scenarios, primarily due to the lack of extensive labeled datasets. One emerging solution is constructing virtual training samples by interpolating inputs and labels, which has notably enhanced system generalization i…