← Search

Wenyong Huang

3 accepted papers

2024

Gaining Wisdom from Setbacks: Aligning Large Language Models via Mistake Analysis

ICLR 2024poster

The rapid development of large language models (LLMs) has not only provided numerous opportunities but also presented significant challenges. This becomes particularly evident when LLMs inadvertently generate harmful or toxic content, either unintentionally or because of intentional inducement. Exis…

Cited by 36SourcePDFScholar
2023

Improving End-to-End Speech Processing by Efficient Text Data Utilization with Latent Synthesis

EMNLP 2023long findings

Training a high performance end-to-end speech (E2E) processing model requires an enormous amount of labeled speech data, especially in the era of data-centric artificial intelligence. However, labeled speech data are usually scarcer and more expensive for collection, compared to textual data. We pro…

Cited by 0SourceScholar
2022

SPIRAL: Self-supervised Perturbation-Invariant Representation Learning for Speech Pre-Training

ICLR 2022poster

We introduce a new approach for speech pre-training named SPIRAL which works by learning denoising representation of perturbed data in a teacher-student framework. Specifically, given a speech utterance, we first feed the utterance to a teacher network to obtain corresponding representation. Then t…