← Search

Liqiang He

5 accepted papers

2024

Generative Pre-trained Speech Language Model with Efficient Hierarchical Transformer

ACL 2024long

While recent advancements in speech language models have achieved significant progress, they face remarkable challenges in modeling the long acoustic sequences of neural audio codecs. In this paper, we introduce Generative Pre-trained Speech Transformer (GPST), a hierarchical transformer designed fo…

2023

Bidirectional Alignment for Domain Adaptive Detection with Transformers

ICCV 2023poster

We propose a Bidirectional Alignment for domain adaptive Detection with Transformers (BiADT) to improve cross domain object detection performance. Existing adversarial learning based methods use gradient reverse layer (GRL) to reduce the domain gap between the source and target domains in feature re…

Cited by 17PDFcodeScholar
2022

DP-DWA: Dual-Path Dynamic Weight Attention Network With Streaming Dfsmn-San For Automatic Speech Recognition

ICASSP 2022accepted

In multi-channel far-field automatic speech recognition (ASR) scenarios, distortion is introduced when the speech signal is processed by the front end, which damages the recognition performance for the ASR tasks. In this paper, we propose a dual-path network for the far-field acoustic model, which u…

Cited by 0SourceScholar
2021

Learned Transferable Architectures Can Surpass Hand-Designed Architectures for Large Scale Speech Recognition

ICASSP 2021accepted

In this paper, we explore the neural architecture search (NAS) for automatic speech recognition (ASR) systems. We conduct the architecture search on the small proxy dataset, and then evaluate the network, constructed from the searched architecture, on the large dataset. Specially, we propose a revis…

Cited by 0SourceScholar