- Stable Audio: https://stability-ai.github.io/stable-audio-demo/ https://www.stableaudio.com/
- MusicGen: https://ai.honu.io/papers/musicgen/
- MusicLM: https://google-research.github.io/seanet/musiclm/examples/
- Stable Audio: https://stability-ai.github.io/stable-audio-demo/ https://www.stableaudio.com/
- MusicGen: https://ai.honu.io/papers/musicgen/
- MusicLM: https://google-research.github.io/seanet/musiclm/examples/
audio quality is still WIP and UI is ugly but prompting should allow more detail.
e.g., 70s Dolly Parton-style song rando made earlier today: https://sonauto.ai/songs/t4XGcHqsYwS6B73MjtV5
It seems designed for making pop music no one will listen to.
I have spent many hours with MusicLM making wild experimental music no one will listen to.
MusicLM has no problem making really weird sound combinations.
I just gave SunoAI some of my MusicLM prompts I have saved and the results are garbage. The problem with the AI test kitchen model though is the results sound like they are in mono.
The ultimate for me will be when we can make rap/hiphop no one will listen to.
But the best so far is Suno.ai ( https://app.suno.ai ) especially with their V3 model they have very impressive results, the fidelity is not studio quality but they're getting very close.
It's very likely based on their TTS model they have released before (Bark), but trained on more data and with higher resolution.