Noise2Music: Generating Music from Text Using Diffusion Models
noise2music.github.io
noise2music.github.io
I would also propose taking it in the direction of generating synthesizer parameters for a popular VST or Hardware synth instrument. As a musician, it would be very nice to be able to program a synthesizer through plain text as a starting point.
The Spectrogram Model for #23 prompt "It sounds energetic and like something you would hear in clubs." sounds almost EXACTLY like "Psy - Gangnam Style"...
The model is hallucinating what it was trained on.
Guilty of terribly trite, cliched and overall bad music, but not guilty of plagiarism.
My fav is the "hippie coffee shop" jam band clip. That will surely corner the market for Jam band background Muzak at "hippie coffee shops". Total available market of like $5.
At best this new synthesis technique will be an Autechre album.
I don't exactly know how to interpret this prompt, and the resulting solo drums meander around as though they don't either. Not really on threes or waltz or 1/3 notes, but a brief tour through all of these and other rhythms.
but Riffusion[1] uses the spectrogram approach and kind of works.