← Search

Rachel M Bittner

7 accepted papers

2024

LLark: A Multimodal Instruction-Following Language Model for Music

ICML 2024poster

Music has a unique and complex structure which is challenging for both expert humans and existing AI systems to understand, and presents unique challenges relative to other forms of audio. We present LLark, an instruction-tuned multimodal model for *music* understanding. We detail our process for da…

2022

A Lightweight Instrument-Agnostic Model for Polyphonic Note Transcription and Multipitch Estimation

ICASSP 2022accepted

Automatic Music Transcription (AMT) has been recognized as a key enabling technology with a wide range of applications. Given the task’s complexity, best results have typically been reported for systems focusing on specific settings, e.g. instrument-specific systems tend to yield improved results ov…

Cited by 0SourceScholar
2017

Towards the characterization of singing styles in world music

ICASSP 2017accepted

In this paper we focus on the characterization of singing styles in world music.We develop a set of contour features capturing pitch structure and melodic embellishments.Using these features we train a binary classifier to distinguish vocal from non-vocal contours and learn a dictionary of singing s…

Cited by 0SourceScholar
2015

Kernel Additive Modeling for interference reduction in multi-channel music recordings

ICASSP 2015accepted

When recording a live musical performance, the different voices, such as the instrument groups or soloists of an orchestra, are typically recorded in the same room simultaneously, with at least one microphone assigned to each voice. However, it is difficult to acoustically shield the microphones. In…

Cited by 0SourceScholar