← Search

Julian D. Parker

3 accepted papers

2025

Scaling Transformers for Low-Bitrate High-Quality Speech Coding

ICLR 2025poster

The tokenization of audio with neural audio codec models is a vital part of modern AI pipelines for the generation or understanding of speech, alone or in a multimodal context. Traditionally such tokenization models have concentrated on low parameter-count architectures using only components with st…

2024

STEMGEN: A Music Generation Model That Listens

ICASSP 2024accepted

End-to-end generation of musical audio using deep learning techniques has seen an explosion of activity recently. However, most models concentrate on generating fully mixed music in response to abstract conditioning information. In this work, we present an alternative paradigm for producing music ge…

Cited by 0SourceScholar