← Search

Meinard Müller

24 accepted papers

2025

Dense-Sparse Dynamic Time Warping for Customizing Piano Concerto Accompaniments

ICASSP 2025accepted

In this study, we explore how pianists can customize Music Minus One (MMO) concerto accompaniments to match their playing style. Bypassing the need for a symbolic score, often not available digitally, we use three types of audio data: solo piano recordings, MMO orchestra-only recordings, and mixed r…

Cited by 0SourceScholar
2024

Performance Conditioning for Diffusion-Based Multi-Instrument Music Synthesis

ICASSP 2024accepted

Generating multi-instrument music from symbolic music representations is an important task in Music Information Retrieval (MIR). A central but still largely unsolved problem in this context is musically and acoustically informed control in the generation process. As the main contribution of this wor…

Cited by 8SourceScholar
2023

Evaluating Speech-Phoneme Alignment and its Impact on Neural Text-To-Speech Synthesis

ICASSP 2023accepted

In recent years, the quality of text-to-speech (TTS) synthesis vastly improved due to deep-learning techniques, with parallel architectures, in particular, providing excellent synthesis quality at fast inference. Training these models usually requires speech recordings, corresponding phoneme-level t…

Cited by 0SourceScholar
2022

Hierarchical Classification of Singing Activity, Gender, and Type in Complex Music Recordings

ICASSP 2022accepted

Traditionally, work on singing voice detection has focused on identifying singing activity in music recordings. In this work, our aim is to extend this task towards simultaneously detecting the presence of singing voice as well as determining singer gender and voice type. We describe and compare fou…

Cited by 5SourceScholar
2021

Reliability Assessment of Singing Voice F0-Estimates Using Multiple Algorithms

ICASSP 2021accepted

Over the last decades, various conceptually different approaches for fundamental frequency (F0) estimation in monophonic audio recordings have been developed. The algorithms’ performances vary depending on the acoustical and musical properties of the input audio signal. A common strategy to assess t…

Cited by 0SourceScholar
2020

Local Key Estimation In Classical Music Recordings: A Cross-Version Study on Schubert's Winterreise

ICASSP 2020accepted

While global key and chord estimation for both popular and classical music recordings have received a lot of attention, little research has been devoted to estimating the local key for classical music. In this work, we approach local key estimation on a unique cross-version dataset comprising nine p…

Cited by 0SourceScholar
2019

Evaluating Salience Representations for Cross-modal Retrieval of Western Classical Music Recordings

ICASSP 2019accepted

In this paper, we consider a cross-modal retrieval scenario of Western classical music. Given a short monophonic musical theme in symbolic notation as query, the objective is to find relevant audio recordings in a database. A major challenge of this retrieval task is the possible difference in the d…

Cited by 0SourceScholar
2019

Fundamental Frequency Contour Classification: A Comparison between Hand-crafted and CNN-based Features

ICASSP 2019accepted

In this paper, we evaluate hand-crafted features as well as features learned from data using a convolutional neural network (CNN) for different fundamental frequency classification tasks. We compare classification based on full (variable-length) contours and classification based on fixed-sized subco…

Cited by 0SourceScholar
2018

Unifying Local and Global Methods for Harmonic-Percussive Source Separation

ICASSP 2018accepted

This paper addresses the separation of drums from music recordings, a task closely related to harmonic-percussive source separation (HPSS). In previous works, two families of algorithms have been prominently applied to this problem. They are based either on local filtering and diffusion schemes, or…

Cited by 0SourceScholar
2017

Data-driven solo voice enhancement for jazz music retrieval

ICASSP 2017accepted

Retrieving short monophonic queries in music recordings is a challenging research problem in Music Information Retrieval (MIR). In jazz music, given a solo transcription, one retrieval task is to find the corresponding (potentially polyphonic) recording in a music collection. Many conventional syste…

Cited by 0SourceScholar
2016

Harmonic-percussive-residual sound separation using the structure tensor on spectrograms

ICASSP 2016accepted

Harmonic-percussive-residual (HPR) sound separation is a useful preprocessing tool for applications such as pitched instrument transcription or rhythm extraction. Recent methods rely on the observation that in a spectrogram representation, harmonic sounds lead to horizontal structures and percussive…

Cited by 0SourceScholar
2016

Retrieving audio recordings using musical themes

ICASSP 2016accepted

In 1948, Barlow and Morgenstern released a collection of about 10,000 themes of well-known instrumental pieces from the corpus of Western Classical music [1]. These monophonic themes (usually four bars long) are often the most memorable parts of a piece of music. In this paper, we report on a system…

Cited by 0SourceScholar
2016

Triple-based analysis of music alignments without the need of ground-truth annotations

ICASSP 2016accepted

The goal of music alignment methods is to temporally align different versions of the same piece of music. These methods are typically evaluated by comparing the computed alignments to given ground-truth annotations. Creating such annotations is usually very labor intensive. For many musical pieces,…

Cited by 9SourceScholar
2015

Extracting singing voice from music recordings by cascading audio decomposition techniques

ICASSP 2015accepted

The problem of extracting singing voice from music recordings has received increasing research interest in recent years. Many proposed decomposition techniques are based on one of the following two strategies. The first approach is to directly decompose a given music recording into one component for…

Cited by 0SourceScholar
2015

Kernel Additive Modeling for interference reduction in multi-channel music recordings

ICASSP 2015accepted

When recording a live musical performance, the different voices, such as the instrument groups or soloists of an orchestra, are typically recorded in the same room simultaneously, with at least one microphone assigned to each voice. However, it is difficult to acoustically shield the microphones. In…

Cited by 0SourceScholar
2015

Matching Musical Themes based on noisy OCR and OMR input

ICASSP 2015accepted

In the year 1948, Barlow and Morgenstern published the book “A Dictionary of Musical Themes”, which contains 9803 important musical themes from the Western classical music literature. In this paper, we deal with the problem of automatically matching these themes to other digitally available sources.…

Cited by 0SourceScholar
2015

Novel audio features for capturing tempo salience in music recordings

ICASSP 2015accepted

In music compositions, certain parts may be played in an improvisational style with a rather vague notion of tempo, while other parts are characterized by having a clearly perceivable tempo. Based on this observation, we introduce in this paper some novel audio features for capturing tempo-related i…

Cited by 0SourceScholar