arXiv · 1604.08723
Music transcription modelling and composition using deep learning
Abstract
We apply deep learning methods, specifically long short-term memory (LSTM) networks, to music transcription modelling and composition. We build and train LSTM networks using approximately 23,000 music transcriptions expressed with a high-level vocabulary (ABC notation), and use them to generate new transcriptions. Our practical aim is to create music transcription models useful in particular contexts of music composition. We present results from three perspectives: 1) at the population level, comparing descriptive statistics of the set of training transcriptions and generated transcriptions; 2) at the individual level, examining how a generated transcription reflects the conventions of a music practice in the training transcriptions (Celtic folk); 3) at the application level, using the system for idea generation in music composition. We make our datasets, software and sound examples open and available: \url{https://github.com/IraKorshunova/folk-rnn}.
Explore related subjects
Keep this discovery
Bob L. Sturm, João Felipe Santos, Oded Ben-Tal, Iryna Korshunova. 2016-04-29. Music transcription modelling and composition using deep learning. https://arxiv.org/abs/1604.08723
Cite the original work for its findings. Save a collection to share your selection of sources.