In plain words: A learning model studies one example song and copies its structure, melody, chords, and bass to make new music in that style. Tested on 10 pop songs, it made high-quality tunes that sounded close to the original, unlike models trained on huge genre collections.
Abstract
Many practices have been presented in music generation recently. While stylistic music generation using deep learning techniques has became the main stream, these models still struggle to generate music with high musicality, different levels of music structure, and controllability. In addition, more application scenarios such as music therapy require imitating more specific musical styles from a few given music examples, rather than capturing the overall genre style of a large data corpus. To address requirements that challenge current deep learning methods, we propose a statistical machine learning model that is able to capture and imitate the structure, melody, chord, and bass style from a given example seed song. An evaluation using 10 pop songs shows that our new representations and methods are able to create high-quality stylistic music that is similar to a given input song. We also discuss potential uses of our approach in music evaluation and music therapy.
Shuqi Dai, Xichu Ma, Ye Wang, Roger B. Dannenberg
arXiv:2105.04709 · cs.SD, cs.AI, eess.AS · submitted May 10, 2021
abstract · pdf · html · 26 pages, 12 figures