I don't mean this as a slight at all, but definitely raise the bar on your experiments.
I don't mean this as a slight at all, but definitely raise the bar on your experiments.
But then I listened to the original (the track used to train the network) and realized the problem: the network only knows how to write one song. What you hear on SoundCloud is the equivalent of giving someone a 5 paragraph essay, and then telling them to write a 10,000 word paper using only sentences contained in that essay.
Supposing that this program can accept more than 1 song in its training data, I expect it could produce really interesting stuff.
But there's a part of music where human soul needs to be, and that is interesting too, and some of the expression stuff is harder to do in MIDI land, you can modulate a filter cutoff or velocity or something - but compared to a live player there is a LOT of work to do.