Generating Music with a Self-Correcting Non-Chronological Autoregressive Model
arXiv:2008.08927
Abstract
We describe a novel approach for generating music using a self-correcting, non-chronological, autoregressive model. We represent music as a sequence of edit events, each of which denotes either the addition or removal of a note---even a note previously generated by the model. During inference, we generate one edit event at a time using direct ancestral sampling. Our approach allows the model to fix previous mistakes such as incorrectly sampled notes and prevent accumulation of errors which autoregressive models are prone to have. Another benefit is a finer, note-by-note control during human and AI collaborative composition. We show through quantitative metrics and human survey evaluation that our approach generates better results than orderless NADE and Gibbs sampling approaches.
8 pages, 4 figures
References in corpus (7)
- Generating Long Sequences with Sparse Transformers
- C-RNN-GAN: Continuous recurrent neural networks with adversarial training
- Professor Forcing: A New Algorithm for Training Recurrent Networks
- Insertion Transformer: Flexible Sequence Generation via Insertion Operations
- Counterpoint by Convolution
- KERMIT: Generative Insertion-Based Modeling for Sequences
- The Bach Doodle: Approachable music composition with machine learning at scale