← Search

Ming-Tso Chen

3 accepted papers

2019

CNN Based Two-stage Multi-resolution End-to-end Model for Singing Melody Extraction

ICASSP 2019accepted

Inspired by human hearing perception, we propose a two-stage multi-resolution end-to-end model for singing melody extraction in this paper. The convolutional neural network (CNN) is the core of the proposed model to generate multi-resolution representations. The 1-D and 2-D multi-resolution analysis…

Cited by 0SourceScholar
2018

A Hybrid Neural Network Based on the Duplex Model of Pitch Perception for Singing Melody Extraction

ICASSP 2018accepted

In this paper, we build up a hybrid neural network (NN) for singing melody extraction from polyphonic music by imitating human pitch perception. For human hearing, there are two pitch perception models, the spectral model and the temporal model, in accordance with whether harmonics are resolved or n…

Cited by 19SourceScholar