← Search

Dejing Xu

2 accepted papers

2022

Audio-To-Symbolic Arrangement Via Cross-Modal Music Representation Learning

ICASSP 2022accepted

Could we automatically derive the score of a piano accompaniment based on the audio of a pop song? This is the audio-to-symbolic arrangement problem we tackle in this paper. A good arrangement model should not only consider the audio content but also have prior knowledge of piano composition (so tha…

Cited by 0SourceScholar
2019

Self-Supervised Spatiotemporal Learning via Video Clip Order Prediction

CVPR 2019poster

We propose a self-supervised spatiotemporal learning technique which leverages the chronological order of videos. Our method can learn the spatiotemporal representation of the video by predicting the order of shuffled clips from the video. The category of the video is not required, which gives our t…

Cited by 563PDFScholar