AI-generated sad girl with piano performs the text of the MIT License
twitter.com
twitter.com
This is literally some of the impossible sci-fi tech I dreamt of as an undergrad. Crazy. I'm still a bit in disbelief how fast things currently move on this front.
Interestingly, suno.ai is also able to imitate the very robotic and staccato-like intonation of Vocaloids: https://app.suno.ai/song/f43e9c46-92d3-4171-bdd9-026213d6772... - everything comes around. :)
Unironically very good. A convincing replica of Miku's voice. Plus the beat itself is great too. As another commenter put it, it is indeed, a banger.
Kaibutsu: https://www.youtube.com/watch?v=-5M4lbEpn6c
The Fox's Wedding: https://www.youtube.com/watch?v=khNi_6PnvaE
What a banger.
* sublicence - "sublissence"
* fitness - "fisted"
* infringement - "infring-ment"
* liable - "liar-ful"
It's also obviously not a pure human voice recording as the pitch transitions sound heavily auto-tuned or electrified (think Cher's "Believe").
I anticipate people becoming experts in detecting AI-generated vocalists in much the same way that we can currently detect AI-generated images due to abnormalities especially in details like ears or fingers.
I am certain there exist weird wines that could fool me (I’ve had a few really weird wines) but typical shit from the grocery store, I’m gonna be able to tell at least that much. I might even ID them more precisely than red or white. It’s not exactly subtle…
Then again I don’t have a clue how someone could fail to tell which is coke and which Pepsi in the “Pepsi challenge”. They’re wildly different flavors. I can tell by smell alone.
So it's less a case of "they cannot distinguish red from white" and more a case of "they went along with a suggested classification". I feel like this is a weaker result, although it's still a little surprising.
Quick, label all the US states: https://imgs.xkcd.com/comics/label_the_states.png
I've given this map to half a dozen smart/well educated Canadians, who happily engaged in pointing out the states they recognized for several minutes, and not one of them noticed until it was pointed out.
https://www.explainxkcd.com/wiki/index.php/2868:_Label_the_S...
For comparison, imagine someone showing a piece of Picasso to art critics and saying "Could you please describe the artistic significance of this painting by da Vinci?" The critics won't start using terms commonly reserved for Renaissance era; they'll say "What the fuck are you talking about, this isn't da Vinci."
The hard part is identifying the type of wine, but many of my wine-drinking friends can do with ease. We've tried the "test," having me or someone else randomly purchase wines from the closest store and then serving random samples to them while they're blindfolded. They're able to identify the specific variety more than 4/5 of the time.
Source: my wife's godfather did the studies for that[1] two years ago.
[1]: https://www.isvv.u-bordeaux.fr/fr/diplome-universitaire-dapt...
What’s a good place to find out the SOTA of the custom tooling and workflow?
Right. Full of code injection vulnerabilities.
"Quadrupled" is a very specific and quantitative word. What measure are you basing that on?
https://soundcloud.com/rs-539916550/soul-of-the-machine?utm_...
IMHO many of the successes of "artificial intelligence" come from "natural stupidity". Humans have many glitches in our perceptual mechanisms. The AIs that end up going viral and become commercially viable tend to exploit those perceptual glitches, simply because that's what makes them appeal to people.
I just rewrote the lyrics as "paste-ing" and it sung it perfectly afterwards.
People fail to identify even the most basic and obvious fakes, but somehow there’s a group of people who think that as fakes become harder to distinguish from reality, we’ll all magically become experts at it. We won’t. People’s ability to detect fakes will get worse, not better, as a consequence of more prevalent and better fakes.
I didn't mean to imply that everyone will become experts; I meant that some people will become very good at distinguishing AI vocals.
Until it becomes impossible which seems to be the trajectory here, no?
The cake is a lie, but the music is real
It’s all fake when the truth is revealed
> (think Cher's "Believe")Or think GLaDOS. Pretty sure that's not a coincidence.
(like it's sort of no difference than paying someone to voice something and share it)
I think the stuff that is completely generated with no human in the loop is a different category for me because it can be used for things at scale like, bots on social media, or ads in a podcast generated just for you, etc. As long as there is still a human in the loop making the editing decisions, it feels not categorically different from the world we have today.
AI/Human combos can still be valuable. More broadly I'd argue that that's how almost all tech works. E.g. there are still textile workers, just many less of them producing much more clothing.
Like, there's been computer-singing voices for awhile, but they always sounded pretty robotic and goofy (e.g. Microsoft Sam), and I think for a long time people just assumed that to get mostly-realistic voices, you need an actual singer. Yes, it still requires a bit of human tweaking to make it perfect, but I suspect that if put to the test it would reduce the cost of making a song substantially.
Free accounts are queued so it depends on load and I don't think the v3 model is available to them.
A lot of people are very confident about this and I dont understand why. The same was said for jazz and comic books. But I am listening to jazz with comic book posters on my wall. There were different reasons to give the same statement, but it almost always turns out to be wrong. Humans like what they like and seldom judge an artwork for its process (outside of a very small niche community).
I see no reason that someone wouldn't hang an AI poster print in their AirBnB that's abstract, or even a "commission" based on the property and its locale.
Same reason for AI music. Your store needs a bit of music to set a vibe, but with AI it can be free of copyright and performance licensing, and again, be tailored to your location, products, clientele, the day of the week, etc.
>Humans like what they like and seldom judge an artwork for its process (outside of a very small niche community).
That's true, but how do you zoom out of process? This is beyond process. I would just say most people don't like inhuman things.
Can you elaborate on this a bit? Because this is what I don't get.
It's an omellete. There is no Dolly Parton behind an AI Jolene or a Michael Jackson turning a 4 track tape into a musical masterpiece. The journey and personalities are what contextualize the sound, without AI that context is gone. That's why I think it will just be used for cafes and things like that where they want to escape licensing fees.
As for consumers - I believe people will see AI music consumption as a way of supporting the new technological powers that be, and the act of listening to human-made music will have an element of counter-culture baked into it. I'm a professional musician and I have a very physical reaction to sound. Once I know it's AI my goosebumps fade.
Another lame incarnation of a tech that will also fade like crypto and everything else. The types of personalities who will leverage this tech are not the same personalities that make the greats.
I'm not worried.
And anyway, is it measurably different from art produced while tripping on LSD or in similar states of altered consciousness, such as schizophrenia, dementia, or even depression, which often produce things many people would not describe as regular?
Counter-point, we've been listening to a rock song about the moon given the words from one of my kids books all morning.
People will 100% listen to (edit - I never finished the sentence) things that sound fun. It might not bring me to tears or stop me in my tracks but lots of things are just fun.
> People aren't going to hang AI paintings in their houses
People will absolutely do this. If AI systems can make nice pictures people will hang them in their houses. And they can make nice pictures.
The future is obviously a form of custom AI Muzak/Ambient music with a few pop stars for people to focus on.
I am a big fan of more art type music and guess what? No one listens to it. My fav album of 2023 has 6.4k views on youtube. At least a 100 of those are mine. No one listens to this stuff. People watch video critic reviews of more art type music than the actual music itself.
https://app.suno.ai/song/2a5a9327-5b27-4353-b62b-8eb3e314fff...
I normally do noisey acoustic stream on consious stuff but I've been too busy for the past two years to get anything out.
This was the first track in a while I've been happy with the lyrics are very real to me and it took a couple hours to learn the workflow and I still haven't went back and fixed edited the ending as well as adding overdubs throughout as well as a real guitar solo.
I can post my youtube for more context but I didn't feel right posting it without someone requesting.
I'm just very interested in new ways to get my art out. And this is more of a transformer of my poetry into real listenable music.
I'm very excited.
Even if the AI music is extremely good, it's just missing the fact that it was made by a person, which changes the experience entirely. I think we're more likely to see musicians and those top10 artists leverage AI without explicitly saying so.
I expect we will have a daft punk moment where someone is using exclusively AI and later unmasks that it was all AI, and as soon as that happens the music is disconnected.
Same with AI art. I can see something and be duped and go "oh wow!!!" and as soon as I know it's AI the caring leaves my body completely and reverence and interest is lost.
It's better than AI, even this incredible mindblowing suno thing. Production value counts.
There’s certainly a point where this synthetic music gets good enough to replace the elevator music Muzak crap that they have to pay $2000 to license.
I need to stem it, fix it up a bit, and remix for stereo in a DAW but it's much better than I expected for my first ever piece of music. Obviously it'd take a lot of work to create a Hans Zimmer level OST from the tool but IMO it wouldn't feel out of place on a Ludovico Einaudi album or on some Spotify or Pandora classical radio.
I don't think musicians and composers are going to disappear as a consequence of this technology, in the same way that theatre actors were not made obsolete by film. What I do think is that a whole new category of professionals will be created - musicians and composers who get paid to train AI models. I bet it will pay better then the laughable amounts that are streaming royalties.
Pretty soon we'll be reviving the old Palmolive "You're soaking in it" commercial (https://www.youtube.com/watch?v=_bEkq7JCbik).
We'll all be soaking in it, and no, you won't be able to tell the difference.
https://app.suno.ai/song/13cffa0c-bbd5-41b6-abde-43332b21b0f...
I took the litany of fear from Dune and got Bing Chat to re-write it to be about facing down code complexity, then I put those lyrics into suno.ai to turn it into a 2 minute song to express all your emotions about code that needs to be simplified ;-)
https://app.suno.ai/song/7995f966-6265-4b34-a68e-400981f5931...
https://app.suno.ai/song/40d0fb88-246b-42f9-8998-0387e75262e...
When they allowed longer text inputs, and for faster rapping, I can really see this kinda thing taking off with L1s and med students.
Like the Animaniacs song about the state capitols.
Or like a Homeric epic that is meant for remembering and singing.
The method of loci may have a new competitor as a way to remember things here.
https://www.youtube.com/watch?v=GVxJJ2DBPiQ
Diagnosis, Wenckebach *(what?)*
It's AV nodal block and that's a fact *(yeah)*
Take PR interval and lengthen that *(yeah)*
bradyarrhythmia and heart attack *(oh-no!)*
AI songs do make sense if AI will be making the diagnosis!https://app.suno.ai/song/018ca476-803b-4a45-81ed-c7263e08ef3...
[1] Ha, the poor millions of dumb minions who put their work on the web thinking it might be fun for others or garner themselves a small following, they didn't check the terms of the EULA!
"Write a sad song about the MIT license" is certainly such new input, and if I was commissioned to write the song it would be based on inspiration (i.e., "use training on") music I have heard or studied. And yes, none of the musicians I have listened to or have studied will benefit from the endless money fountain I'd acquire from composing such song.
In the case of AI, a person puts in minimal effort to generate something that devalues the work of all the people who did put in effort.
When someone needs something composed, they don't learn how to write music. They pay someone else the bare minimum, e.g. a few bucks on fiverr. The person will spend the least possible amount of effort to try to make their life go around with the little money they got.
When you then use an AI model, the work done for those five bucks is replaced by work done for almost free.
Neither the person you would hire or the AI credited those who created the material they trained on.
In other words, you pay a few cents to big tech for a generator that only exists thanks to work real composers, singers etc who now get the grand total of 0$.
We also stopped hiring computers (the occupation) and instead pay big tech companies which made computers (the device) available to everyone. And we stopped hiring people to do dangerous manual labor as companies started selling machinery to automate it. Markets change.
Yes.
> computers
> dangerous manual labor
If people are not in danger and don't have to do mechanical work, it's one thing. If composers stop composing original music that was used for training current AI because they don't get paid anymore then the field stagnates. Same for writing and everything else
There are now far more options, both on the high and low-end, with the whole area being more affordable. The quality of most products also arguably went up, as factories beat handmade goods. And yet, if you want custom artisan goods, you can still pay a woodworker for it at more or less the same cost as you would have otherwise, as their labor costs are a function of time required and local living conditions.
In some cases, those workers were the ones to automate, benefiting from the assistance - woodworkers using CNC mills and laser cutters even for handmade goods, or composers themselves can use the AI - to speed up their otherwise fully manual work. It benefits the majority creating the demand, and tends to improve the craft overall.
> ... then the field stagnates.
A market that is not changing has already stagnated.
> A market that is not changing has already stagnated.
Yep. The way art is changing is thanks to original work and no one will be making it since anything you make gets stolen for free
Speak for yourself! There is only one thing that scares me more than composing music, and that’s paying somebody a few bucks in fiverr to do it for me.
Although I suppose royalty-free stock music is the norm nowadays for most commercial uses, which takes it a step further, anonymizing the composer entirely...
That's by choice though?
And that's the point: The difference is the replacement of 1 flesh-and-blood composer with 1 virtual composer, with the consequence being the lost business of the former. The artists studied were never part of the transaction in either case.
Now, the long-term consequences for artists - e.g., reduced supply on the low end as they're out-competed - is harder to guess, but that's just market dynamics. It may very well increase supply as composition becomes more available, diversifying by allowing people with other skills or creative treats to create music that previously could not - even if the musical part is done by AI.
but with AI whatever consequences there may be their work is highjacked/stolen.
You can learn yourself but if you use an automatic tool to bypass and automatically make similar works and compete with original authors then you're IP thief
Worded differently: people who couldn't otherwise produce skill-based works of value have had the barrier of entry lowered for that specific medium of expression, allowing for more works across a wider spectrum of skill.
Flooding the world with unpolished, unpracticed works, AI-tuned to the level of being mediocre, is a creative and intellectual dead end.
This is what tools are.
Cheap digital tablets have done away with the need for expensive consumables. You can just download a different brush style instead of learning a physical technique. No waiting for paint to dry or smudged pencils. The barrier to entry for painting has dropped to a one time investment of like a hundred bucks. Almost nobody mixes their own paint, nor stretches their own canvas. Those skills aren't needed anymore.
It's possible to build very precise machine parts by hand. It's very difficult and requires great skill, so nobody does that. Some do and are admired for it, but everybody else uses precise machines to make precise parts with nearly no effort.
It's just a tool. Only difference is that we had assumed art would never be automatable.
Objectively, I don't think this is a bad thing. It doesn't change the subjective value of art any more than the average cartoonist devalues the Mona Lisa. It's just a new form of art, there will always be people mixing their own paints and stretching their own canvas, just as there always has been.
It's only a problem because in our society you either have a job or you starve. No one can afford to be an artist. Those that do tend to grind out as many pieces as fast as they can so they can pay the goddamn rent. If not for that, these AI tools would be pretty cool.
* That "_any_ form of creative expression" is a viable creative substitute for people wanting to create in a _specific_ medium of creative expression -- especially those that had a high barrier of technical skills required to be seen as "good enough" to share.
* That a person who has an idea for art will put in the necessary time to become proficient enough to create that "good enough" art through traditional means (IMO demonstrably incorrect), and that is preferred over that person just not expressing a lower-quality version of that idea at all.
* That those who use AI primarily want or expect to "perform at the level of the practiced and talented" (i.e. top-tier art) rather than using it to produce art they otherwise couldn't have, even at low- and mid-level qualities.
* That there is no skill or talent in using AI tools to produce art (or that the skill or talent using AI tools is meant to be a full replacement for traditional artistic skills or talents).
FWIW, I'm a long-time sketch artist and acrylics painter (~20 years). There are many mediums, subjects, and styles that I'm not good at -- and I enjoy using AI to express myself in those areas (and have also liked using AI to create songs to show to my more musicially-adept wife...). But even in my own wheelhouse (landscapes and still life), I also often use AI to brainstorm composition, perspective, colors, textures, lighting, etc. It's a great tool for experts to lean on, but an even better tool for non-artists who couldn't or wouldn't otherwise share their art.
Why is this the line? Where are the complaints about people using pianos to achieve rather precise notes instead of using their own voices? They are just untalented at singing and their use of any tool to create sound is of no net gain to anyone.
This person: https://www.youtube.com/watch?v=IbUE-LxhUR8 ? They're recording and playing back on a loop! They should record full repeated playings, any use of the recording is of no net gain because it could be achieved otherwise.
Songwriters? If they write lyrics and someone else sings them the result should be cast into the sea - it's of no net gain to anyone because they did not create the sounds themselves.
Composers? Frankly pointless.
> Flooding the world with unpolished, unpracticed works
I hate to break it to you but there are a vast number of terrible works of art out there already.
> What this tells you is that skill or level of performance is not the barrier, but a means through which great things CAN be achieved (i.e. necessary, but not sufficient)
If it's a necessary thing, of course it's a barrier. That there are two barriers doesn't change that.
That a person can't sound like the weighted average is human limitation (although with modern pop people do get quite close!), not because new singers aren't trying to. That of course adds variation that we appreciate, but doesn't change the underlying similarity in how acquired skill is mimicry of those who acquired it before us - with very rare exceptions.
What next, are we going to argue that what programmers creating new programs are really trying to do is generate a prompt-weighted average of the bytecode of every program they've ever downloaded, and all that business analysis and functional spec and use of high level programming languages and expressed preferences for coding standards is irrelevant?
That's just a bias.
> natural qualities to their voice
That's the physical limitations I referred to, which isn't something humans tend to be happy about but can sometimes end up being a differentiating benefit.
> What next, are we going to argue that what programmers creating new programs are really trying to do is generate a prompt-weighted average of the bytecode of every program they've ever downloaded
That's a horrible strawman. Do you as programmer often read and write bytecode directly?
I'm beginning to assume you're an LLM, because I'm not convinced a human would honestly try to argue that their emotional reaction to their favourite songs is basically equivalent to flipping the values of some bits to ensure that they generate music more similarly to them.
> That's a horrible strawman. Do you as programmer often read and write bytecode directly?
As an improvising guitarist (even a very mediocre one) my creative process is even further removed from an LLM parsing and generating sound files directly....
We are all nothing but a horde of molecular machines. Your "you" is just individual neurons reacting to input in accordance to their current chemical state and active connections. All your experiences, unique personality treats, and creativity you add to the process is solely the result of the current state of your network of neurons in response to a particular input.
But while an LLM is trained once and then has its state fixed in place regardless of input, we "train" continously, and while an LLM might have experience of an inhuman corpus for a certain subject, we have many "irrelevant" experiences to mix things up.
Your "prompt" is also messy, including the current sound of your own heartbeat, the remaining taste in your mouth from your last meal, the feeling of a breeze through your hair as it tickles your neck, while the LLM has just one, maybe two half-assed sentences. This mix of messy experiences and noisy input fuels "creativity". You don't think "I need to copy XYZ", but neither does the AI. You both just react.
In some regards our chaos is better, in others it is worse. But while the machinery of an LLM still does not even remotely approach a brain, we should not forget that we are nothing but more a cluster of small machines, assembled from roughly 750 MB worth of blueprint.
All of the creators and subjects of meme formats... Should they receive royalty every time you post some inane mashup?
I don't begrudge crypto miners either.
This is the real problem, right? People don't dislike generative AI, they dislike the attention economy. Yet I see more disgust towards AI than the company policies which suck. I don't understand why.
Hell, even with viral videos it's relatively common that normal people can share away while entertainment companies and influencers are expected to pay for a license.
With memes it isn't clear exactly who made the first template, and the creation of them doesn't revolve around specific people in the same way, nor are they meaningfully tied to profits.
When creators post their content online to be shared, they do it with the focus being on reaching individuals, not for it to be sucked up by soulless companies to extract all value without the intention of giving back.
The Office, The Matrix, Lord of the Rings, Django Unchained, Game of Thrones, etc
These works have identifiable creators.
We're not talking about any of those things. We're talking about wholesale digestion of the entirety of human knowledge by automated means, which is now not just theoretically possible, but routine.
Unless Devin has his way.
Stephen King won't be able to remember every word of every story he's ever read. And if he wants to make something "Lovecraftian", it'll be what Stephen King thinks is Lovecraftian. And there will be something to that. Some bit he believes is more or less important than other people And those bits are what makes Stephen King, Stephen King.
Everyone has had access to the same material King read. Access to the same tools he used to create. Everyone had the chance to effectively be Stephen King. But there is just one. Because there is some unique bit of observation or recall or combination of such things that is unique to King.
And from what I've seen so far, these LLMs can't do that. There is a missing element of pure imagination.
Saying they cannot create based off of a vague suggestion is very much in line with that claim. I consider it a vital difference between Stephen King being inspired and LLMs mashing training inputs together.
Musically, however, I can't help but notice that these models are still very far from being able to generate something interesting: from harmony, to tempo, to musical structure, to dynamics, everything is muddled and without structure. I guess there is still very much to work on, and I am not sure that purely generative models can attain higher levels. Maybe a mixed rule-based and generative approach would do?
The progress is really fast in this field, I really do not know.
The github page for bark links to a page about chirp, which returns a 404 page for me [2]. My guess is that the model used for suno.ai's song generator isn't too much different than the text to speech model.
I also have a hunch is that it was more like a coincidence than intentional that the bark model was capable of producing music, and that was spun off into this product.
Unfortunately, there seems to still be issues with bark when generating long (like book length) spoken audio. Which is too bad, as someone who's worked jobs that require lots of driving, it would be awesome to be able to have any text read to me in a natural sounding voice.
[1]https://github.com/suno-ai/bark [2] https://www.suno.ai/examples/chirp-v1
Elton John improvising on an oven manual is the high bar in my opinion. https://m.youtube.com/watch?v=8GuI4UUZrmw
Music is a language, even if with no semantic. It has conventions, dialects, a syntax, a grammar. There are multiple dimensions a musician uses to convey what he wants/feels: just like an actor has to control at the same time its voice, posture, interplay with other actors, so a good musician is aware of the structure of the piece he is composing/executing, the relations between the various subparts, how the musical discourse progresses in time, besides agogic, dynamics, sound color.
All of those aspects are continually perpetually compared against the conventions of the genre, mixed, evolved, strictly followed or balatantly negated.
This is something that normally a professional musician takes decades to master (apart from musical geniuses).
A listener takes less time to educate himself to appreciate those nuances (but not too little: let's say ~years). Once you develop a taste, it becomes very obvious to see through the spectrum that goes from bad quality tunes to musical artistry.
I see nothing musically interesting in this (wonderful) PoC of speech synthesis.
Just to be clear: I did not see anything particularly stunning even in Google's Bach Doodle from some years ago https://doodles.google/doodle/celebrating-johann-sebastian-b...
I wonder if these models would do something better if the text were poetic or punctuated differently.
I generated lyrics with chatgpt4 + some manual tweaking.
I feel the suno-generated song itself lacks congruency. Like, you could listen to it 10 times, have the lyrics in front of you, and not be able to sing along.
The software is provided (as is)
without warranty, of any kind
express, or implied.https://app.suno.ai/song/2ce5eab5-d1c5-48b2-91a0-8e6095e29ed...
https://www.gnu.org/music/free-software-song.en.html: “Richard Stallman and the Free Software Foundation claim no copyright on this song.”
https://app.suno.ai/song/f283429e-ec3e-4152-b5be-a57cd72a6d9...
They've been listening to it in the bath to huge success. Particularly the change at about 40s for "why can't we breathe on the moon?" which feels like an excellent song lyric.
Honestly I'm blown away at how well it does.
break up chapters or sections of the college textbook into suno songs instead - itd be maad interesting how much better that wouldve helped my studies. monotone computer voices of 10+ years ago will put you to sleep.
I'm not sure whether I've just run out of credit, or Suno actually knows what the political sensitivities of the text might be, but I can't generate a second amendment song.
And some of the generated phenomes actually just sound like stylistic auto-tuning. I kind of like it.
I'm sure many have already observed this, but I think the thing that most artists fear from AI is not that AI will be able to produce works on parr or superior to human works, but that most people won't care enough to value the difference.
I think I'm going to enjoy how surreal widespread access to generative AI will make the world.
https://app.suno.ai/song/d67a4c29-9f2a-41b0-9ff8-2c8138a1a7a...
https://app.suno.ai/song/89a48c01-c7e5-487a-b825-4a3978b7259...
https://app.suno.ai/song/30f8223e-0d0b-4cac-8b3f-5d8f0f743e2...
The lyrics generator is some version of GPT so you can give it natural language instructions like this.
https://app.suno.ai/song/4f96f485-8d84-4df0-9a9c-941984137cc...
SVD doesn't have well-controllable motion and is utterly blown out of the water by Sora, but it's what we have right now.
Here's the resulting vid, "a death metal song about a macro photographer":
https://www.youtube.com/watch?v=kNVRQ1Zg-a0
If you only want a video file from Suno to share with the default static lyrics screen on it, hit Download Video from the three-dots menu.
I've dabbled in music production and this is just unbelievable.
Both amazing and a bit sad because this is already so much better than would i would have anticipated.
First illustrators, copywriters, then VFX guys, and now music. We're going to loose so many jobs in the creative sector right?
(Dual Core FTW!)
Long-term coherence, reasonable-ish melody, all on top of very unmusical text. Very impressive.
Ok, this is pretty fun.
https://simulationcorner.net/SAM/sing.wav
Edit:
Suno, or its next clone, will dramatically close that technological gap; no army of lawyers could stop it. I would brace for impact into a short interim period of dead talent revival. Then, newly synthesized, superhuman artists will emerge.
https://app.suno.ai/song/a693c847-7ce6-475c-adc5-0328786b901...
Haha this is amazing!
Soon we will have 'preacher's in a box' that will sing to lift you up, mentor you, guide you through life. Most will even be 'non-religious' but will basically become your religion, your guide through life.
* Put a poem my late mother wrote to music for her memorial
* <asian-language> versions of 80's new wave songs
and they came out so lovely compared to what I'd be capable of as a musician, but puts me in the role of a "producer" of sorts tuning the sound and vibe. Really well worth the money.
https://app.suno.ai/song/41fde9b6-a722-4c39-92dc-8a8296c018c...
Anyway, what will be interesting is when this can be done locally on consumer hardware with open-source AI, a nice UI and Vulkan/DirectML GPU inference.
Current generated songs are made like sentences where you hear entire song without much structure
Wonder how long it will be until someone sings Mein kampf though
[0]: https://suno-ai.notion.site/FAQs-b72601b96de44e5cacd2cd6baa9...
The cake is a lie, but the music is real
It’s all fake when the truth is revealed
The autotune electronic voice seems likely styled on GLaDOS.the ending is cool and unexpected
I’m making a note here: “Huge success”.
Your version unfortunately sounds much more plausible and profitable.
https://www.x.org/releases/X11R7.6/doc/xorg-docs/specs/ICCCM...
https://app.suno.ai/song/86040709-94f1-4de1-8703-7b306b48b32...
ICCCM Summary of Window Manager Property Types
https://app.suno.ai/song/52d08a23-8e1e-4f03-8e8c-e4df610cef9...
If this is legit, the Spotify spam is going to become atrocious and probably unmanageable.
I had done a folk song version of my resume. It wasn't going to become a hit or anything, so I don't see this replacing any real musicians, but it absolutely worked to create a passable performance as a song.