Learning and Evaluation Methodologies for Polyphonic Music Sequence Prediction With LSTMs

Adrien Ycart,Emmanouil Benetos

doi:10.1109/taslp.2020.2987130

Abstract

Music language models play an important role for various music signal and symbolic music processing tasks, such as music generation, symbolic music classification, or automatic music transcription (AMT). In this article, we investigate Long Short-Term Memory (LSTM) networks for polyphonic music prediction, in the form of binary piano rolls. A preliminary experiment, assessing the influence of the timestep of piano rolls on system performance, highlights the need for more musical evaluation metrics. We introduce a range of metrics, focusing on temporal and harmonic aspects. We propose to combine them into a parametrisable loss to train our network. We then conduct a range of experiments with this new loss, both for polyphonic music prediction (intrinsic evaluation) and using our predictive model as a language model for AMT (extrinsic evaluation). Intrinsic evaluation shows that tuning the behaviour of a model is possible by adjusting loss parameters, with consistent results across timesteps. Extrinsic evaluation shows consistent behaviour across timesteps in terms of precision and recall with respect to the loss parameters, leading to an improvement in AMT performance without changing the complexity of the model. In particular, we show that intrinsic performance (in terms of cross entropy) is not related to extrinsic performance, highlighting the importance of using custom training losses for each specific application. Our model also compares favourably with previously proposed MLMs.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

Learning and Evaluation Methodologies for Polyphonic Music Sequence Prediction With LSTMs

Abstract

Talk to us

Similar Papers

More From: IEEE/ACM Transactions on Audio, Speech, and Language Processing

Lead the way for us

Journal: IEEE/ACM Transactions on Audio, Speech, and Language Processing	Publication Date: Jan 1, 2020
Citations: 41

Similar Papers

An automatic music generation and evaluation method based on transfer learning.
Yi Guo ... Yangcheng Liu
PLOS ONE | VOL. 18
Yi Guo, et. al.Yi Guo ... Yangcheng Liu
10 May 2023
PLOS ONE | VOL. 18

Automatic music transcription: challenges and future directions
Emmanouil Benetos ... Holger Kirchhoff
Journal of Intelligent Information Systems | VOL. 41
Emmanouil Benetos, et. al.Emmanouil Benetos ... Holger Kirchhoff
25 Jul 2013
Journal of Intelligent Information Systems | VOL. 41

Future vector enhanced LSTM language model for LVCSR
Qi Liu ... Yanmin Qian
-
Qi Liu, et. al.Qi Liu ... Yanmin Qian
01 Dec 2017
01 Dec 2017

Low rank modelling for polyphonic music analysis.

-

31 Jul 2020
31 Jul 2020

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

Learning and Evaluation Methodologies for Polyphonic Music Sequence Prediction With LSTMs

Abstract

Talk to us

Similar Papers

More From: IEEE/ACM Transactions on Audio, Speech, and Language Processing