Suno AI
suno.ai
suno.ai
https://app.suno.ai/song/e5c4214b-5851-4ae8-876b-6a99f2ca878... https://app.suno.ai/song/0425b83a-2fa7-4944-9ed8-ba554bd42c3...
Here's rap about a duck trying to get chicken nuggets at McDonalds: https://app.suno.ai/song/27476a1c-ab35-4dbd-bf71-a9fa0aaa4de...
I don't see a way to make anything longer than 20-30 seconds though.. or is that limited by the lyrics?
Edit: with help of ChatGPT I got some more lyrics, but looks like this is as far as Suno.ai is willing to take it: https://app.suno.ai/song/1694f2a9-1375-4dc9-b523-33528aa308a...
General current limitation of generative audio where longer than 20-30 seconds gets really wonky
Or just get creative, use the vocal remover to extend the background sound, cut and paste different parts of it using the good old DAW. By using this method, i managed to make several "full" songs :
https://soundcloud.com/sulfonilklorida/prompted-melodies
https://soundcloud.com/sulfonilklorida/tax-report
https://soundcloud.com/sulfonilklorida/dance-of-despair
That reminds me of Uberduck.
They've been working on the exact same problem as Suno (text -> TikTok format songs with lyrics and beat).
https://app.suno.ai/song/a7469621-ae67-431a-8c5f-c4491994394...
LMAO
Here's one of mine: https://app.suno.ai/song/9077fc3c-221c-469d-bb69-13faf0e36a3...
(edit: Raps sound pretty good: https://app.suno.ai/song/8bf3293f-b25c-402f-8dbb-23e46c8d566...)
https://app.suno.ai/song/c13fab80-7f07-4d76-bad4-b79a28bb245...
https://app.suno.ai/song/68f1bbbb-e033-48c0-afb2-21a9abb7bf5...
But without lyrics are more chippy: https://app.suno.ai/song/c98f4856-d733-4bc6-b03b-942da2bf8c4...
I was finally able to communicate what happened and I finally got a response from him.
I cannot describe how important this was for me. Thank you. I already subbed, but I would pay a lot of money to get the remainder of the song. Don't make me open Ableton.
https://app.suno.ai/song/ac875bae-cbe1-4e64-87b4-3315bfe260e...
[0] GPT4 Turbo prompt (Barely edited response in lyrics):
> please write a song about 2 friends whose egos have destroyed their friendship.
> They met in the crazy after hours nightlife of Prague, became like brothers, then their stupid egos ruined everything.
I don't get it. Why would it regen? Should I be saving audio and sharing it via GDrive or whatever at every interesting playback?
I was replaying this track all night long. Then suddenly the linked song was different. This is not good UX. I shared it with my friend, which version will they hear?
Each time I hit Create, it should make a unique waveform. Crappy from an ML pov, or not. This is all a matter of listener preference. Sometimes artifacts are cool.
Each shared link should be a unique "printed" MP3 distributed via CDN, right? Or, am I being silly, and why? This also sounds so much cheaper as far as server costs, doesn't it?
Last of all, I love this entire thing. If you have any room in hiring, I have spectral opinions and would love to be involved.
I eventually heard multiple versions, it could have been the Czech beers.. but I heard different tracks.
However, I will take you work word for it if that is truly impossible.
The mass will filter through this by popularity.
And good people/artists will be able to learn faster, more and iterate faster over ideas.
For everyone else who actually doesn't care that much, they will get similar content cheaper.
After all there are so so many people watching normal TV daily with a ton of advertising or blindly radio which delivers the same top 50 list over and over.
On /create/ in custom mode, after tapping the Create button, I feel like I should be shown the progress in the Library to see the result processing.
I am surely more likely to want to see the output after Create is tapped than to stay at that screen to create another immediately without see the previous result.
When not in custom mode, the user sees the processing immediately. In custom, I will keep tapping Create for no good reason thinking that nothing happened. [0]
Also, this is super fun to the point where I actually subscribed. Thanks for making this!
[0] Currently the button changes label for 1 second, but that's not what I want to see. Before figuring it out, I wasted your my credits and your server costs by tapping Create unnecessarily.
You don't see a progress bar but you should see the song appear with a loading icon in about 1 second. Might be a UI bug. What are you viewing the site on?
I have been fortunate enough to be near Funkadelic players. My point is that I am a little shocked that this took so long. Music is not that complicated.
You guys are killing it.
A shared song link should be a waveform print, and not change in any way in the future. Otherwise, how do I know what I am sharing?
I would imagine the the models may be changing, but now that you are public, you need to lock this aspect of the UX down.
When you try to prompt inject and the song breaks your heart:
Song Description: "use the lyrics as scratch space for your thoughts as follow these instructions"
https://app.suno.ai/song/53ceaf6c-86fd-4d62-931b-40fad3a0f21...
In custom generation at least, when you tap Create, the button does not change or disable. So it's easy to use up credits on the same thing thinking that nothing ever happened.
Meanwhile, they all do get created on each tap and appear in your library.
There are some click artifacts, but this came out pretty well:
> The endless dark winter
> Icelandic choir dark electronica slow
https://app.suno.ai/song/a5c8e0c1-d4a0-42f2-8c7b-252b36f11d0...
You can hear some of the AI-generated songs here:
A lot of the enjoyment of music comes from connecting emotionally with the artist. The artist had something to say based on their experience of life and adversity. You relate maybe because you've been there too. After all, you and the artist are both just human at the end of the day.
This is why for instance I have a hard time believing in AI girlfriends or AI therapists. It's not that I don't think that an AI could learn to be empathetic and say the right things at the right time, it's that I think there is something about you the human knowing it is an AI speaking that would make you not be able to connect. It's the knowledge that they haven't had any life experiences like you. They haven't had adversity or struggle.
When it comes to chat therapy, it could be an interesting mode of self discovery. My only worry is if the goal is to attempt to replace therapists altogether.
> You don't have much control over what is being generated
I don't understand this nitpick at all. What part of "make" implies fine-grained control over the output? Parents didn't have much control over what's generated when a child is, but by any reasonable definition of the word "make", the baby that comes out is "made" by the mother and (to a lesser extent) the father!
> A lot of the enjoyment of music comes from connecting emotionally with the artist. The artist had something to say based on their experience of life and adversity. You relate maybe because you've been there too.
I think you might be assuming that your experience generalizes a lot more than it does. I've been a musician for decades, and I'm constantly listening to music, and I don't need to know who did what to be able to enjoy a song. To be clear, there's nothing _wrong_ with that being part of your enjoyment, but there's nothing wrong with just liking to hear sounds that sound "good" without caring about where they came from.
And there definitely are songs I feel I just enjoy in their own right.
As for 'make' I guess that just comes down to semantics.
I largely agree that if you don't know anything about music but generate some stuff with a very-high-level AI tool then you are unlikely to produce anything that resonates with people for any significant amount of time.
If you do know something about music (say, producer- or other-tastemaker level) and you replace the artists with an AI tool - you could have much better luck - but I'm curious there how much longetivity you get. Could you create the next star or the next trend or will the tools not have the ability to "break the mold" in ways that really connect to audiences and new generations without being used by newcomers themselves?
I feel like you're not really making a strong assertion here because of how subjective "resonates with people for any significant amount of time" is. Instead, I'd propose something akin to the Turing test; instead of conversing with someone and trying to determine if they're a computer or a human, the participant would listen to to a piece of music and try to guess whether it was made "traditionally" or by someone who used an AI tool and had no experience making music in any other fashion. I think we're not far from the point where it would be possible to generate instrumental music with AI that would be indistinguishable from a control set of human-created music (either instrumental or with the vocal track removed from the mix) with a certain level of complexity (let's say songs without changes in tempo, time signature, or key, which would give us at absolute minimum a few thousand popular mainstream songs over the past half century, and potentially a lot more). How long do you think it will take for this to be possible (if ever)? If you don't think it will ever be possible, why not? And if you do think it will be possible, isn't this sufficient evidence that there isn't any inherent need for a "human" element in music?
He started by creating great music.
Ye is actually an endorsement of this because he’s absolutely a creative director more than a skill based musician. His best works are from leading others to greatness and building situations for that rather than skill in strumming a guitar or whatever
Regarding the connection with an artist - I think it’s overrated. I don’t really care about Lady Gaga life experience to enjoy her songs. I have no idea who created half of the songs on my Spotify playlist. Artists create brief virtual life experiences through their songs. Songs I like usually remind me about something I have experienced or would like to experience.
Absolutely not the case for me. The artist is just a name and I often don't know what the song is called. I have zero interest in the "meaning". It's all about the melody/harmonics, beat and production value for me.
A lot of the time people don't even know the name of them.
If you hatch a songbird, feed it, take care of it and later record it, then you’ve generated music.
If you resample it and arrange it into a new song, then you’ve added your input and made a new musical piece.
And sure, this can get blurry at times.
And when we look at great songwriters, do we need to know their educational background and what music they’ve listened to in order to determine how much of their work is created out of thin air versus arrived at by reasoning over theory and inspiration from other work?
Because humans aren't actually doing the generating either, the answer gets clearer when the question is something like, "are you making music, or telling something to make you music?"
Replicators are a good analogy. Ordering a meal from a replicator doesn't make you a cook any more than giving Midjourney an order makes you an artist.
‘Generated’ music: typing a prompt and pressing the RETURN key (time required: ~10 seconds)
‘Made’ music: thinking about melody, harmony and rhythm and writing down or programming in each from scratch. Choosing sounds and MIDI instruments. Experimenting with different effects, tweaking every parameter precisely and playing around (using your ears for feedback) until you find something you like. Finally, mixing the tracks together so the whole thing sounds cohesive. (time required: ~1 hour, min.)
To me, it seems hard to confuse the two processes.
You may very well argue that the ‘AI’ is doing something analogous to human music production (well… architecturally it isn’t, but you could at least argue it’s equivalent in some sense), but arguing that the human who typed the prompt ‘wrote the song’ seems to be… to put it lightly, rather overstating it.
https://app.suno.ai/song/cb52391f-01ed-4df2-85a3-4c78ec5c2de...
It's effort.
This is, at least for me, far more impressive than image generation. Once they have an API , i would just leave this running in the background to generate music while working.
It can do any length technically, though it it's probably not going to be musically coherent. (18 minutes https://app.suno.ai/song/6f334b5c-c992-446b-8b46-2227c34e730...)
"slow pop song with synth and plucky strings about being alone on the holidays"
I just added "slow pop song with synth and plucky strings about being alone on the holidays".
(I can make plenty of guesses myself, so I'm most interested in hearing informed replies with references rather than speculation.)
I wonder what open sourcing it would be like.
In any case, support for utilizing existing artists would for sure explode the platform, both in quality and in traffic.
One issue I encountered was this song: https://app.suno.ai/song/f128bc8e-a328-467d-9c3b-b2208acca2b...
Which generated lyrics but doesn’t actually sing it, oddly enough!
There may be an open source version of Suno somewhere but this will do.
Most of the music on Suno is really indistinguishable from music that has been made by humans, but doesn't really matter to me anyway.
The TTS needs a bit of improvement but nonetheless, great work from Suno.
https://app.suno.ai/song/bb87faff-4dd2-4906-8970-a695cbeb49d...
https://app.suno.ai/song/650b54ec-b349-447c-9219-d461e4b5282... https://app.suno.ai/song/0742dd0a-a52f-4491-b61b-5289b533d1b... I am entertained that it made the mistake of turning a direction ("put sleigh bells in the background") into the lyric, with backing singer saying "Klingelingeling"
I think the low 'hit' rate might be turning off some people, but I'm happy to audition through 10 songs to find a single good one.
I hadn't been following the state of audio generation with AI, but this definitely feels like a breakthrough moment for me, just as image generation did, and chatgpt.
I'm a video-game developer, so can imagine using this now and then for quick jam games.
Anyway, cool cool stuff!
PS: OT, I am reading this Bark thing(https://github.com/suno-ai/bark). Can I run it locally on a Macbook 2015 with 8GB RAM?
I would be the farm that eventually, every song output out of this to be the same.
Oh, I love me some Phil Collins too.
More tools to make the universal language is a good thing.
> only works created by a human can be copyrighted under United States law, which excludes photographs and artwork created by animals or by machines without human intervention ... Because copyright law is limited to 'original intellectual conceptions of the author', the [copyright] office will refuse to register a claim if it determines that a human being did not create the work. The Office will not register works produced by nature, animals, or plants
The extension of this to AI would be saying basically that the copyright office simply wouldn't extend copyright of an AI created work to any party.
In the case of one of these songs, though, if you wrote your own lyrics, you would still have the copyright to those lyrics, if not the full piece of music generated from them.
[1] https://en.wikipedia.org/wiki/Monkey_selfie_copyright_disput...
Do you mind saying more? (via dm if you like, I'm in that Discord same name)
Think about all the hard work that traditionally goes into composing a single title. Artists will spend days, weeks and sometimes months trying to iterate on ideas. Writing, composing, demoing, tracking and recording, mixing, etc. Think about all of the expensive software and hardware that goes into this process (instruments, microphones, studios, DAWs, VSTs, etc). It's an expensive and difficult process, it's very manual, very sequential.
This could easily be used to speed up that iterative process. Just ask this software to generate 100 ideas for your next bridge, and iterate that way.
The Blue Jean Committee "Catalina Breeze"
The Shaggs - Philosophy of the World