Show HN: I had some time yesterday so I made a GPT3 podcast to help you sleep
anchor.fm
anchor.fm
1- Don't host on anchor. Podcasting is an open standard. Don't let companies (like Spotify or Apple) take it over. Check https://podcastindex.org/
2- The voice is too mechanical for this to be actually reasonable to listen to at night, potentially could be listenable with AWS Polly Neural voices, it's pretty good.
I'll try Polly, thanks! The current voice annoys me too.
You can use Anchor to generate your RSS feed and host your content while still sharing the RSS URL on a domain you own. So you'd give out a URL like feeds.deepdreams.com/rss, and it would proxy the response from Anchor's RSS feed
I wrote a simple Go cloud function that can proxy your Anchor RSS URL for you:
It's one of those things that's hard to do after you've got a bunch of subscribers, so I'm always glad if I can warn people early in their podcast against getting stuck with their host.
https://deepdreams.stavros.io/
dang, if you see this, could you repoint the URL to the article? Thanks!
I wrote a script later to automate the audio mixing, that took another hour. Now I can generate a ready-to-upload episode with one command, though.
I initially didn't recognize the username because it was all-lowercase, so I'm curious why the rename from StavrosK to stavros?
Really the main concern I would say is that the author doesn't own the domain so they are locked in, but I don't see how this affects listeners.
Do you know if there are any similar quality TTS tools for less technical applications? I mean, where you can just type in the text you want and get an audio file with a high quality voice?
There are free websites like this: https://ttsreader.com/
But TTS is built into browsers as well eg (this it would need some code not much) https://developer.mozilla.org/en-US/docs/Web/API/SpeechSynth...
I use AWS Polly to read HN in the mornings
But podcasts? There has got to be a more interesting use case than that.
Maybe a female voice, a bit quieter (the soundscapes are almost completely silent for me) and maybe add some high-room-size, long decay (5-10, maybe even 20 seconds), wide panned (like 100%) and moderately diffused (maybe 10-20%) reverb to the voice with like 30% mix or so, which would add a very airy tone and help the voice blend in a bit. If the TTS engine has a whisper setting (many do), add just a bit. It'll help thicken the reverb.
That, paired with bass-heavy soundscapes, will create a very nice balance between the low registers and the voice's high registers.
Just a thought. :)
Actually, fuck it:
https://gitlab.com/stavros/deep-dreams
I'll implement your suggestions (or as many as I can), thanks!
"AAAAAAAh WOOAH Jeez guys, I sure wanted to tell you this AWESOME-SPOOKY story that I had, but I can't read the next word on my sheet because my flashlight broke. ffffffffff! I hate it when it breaks."
How did that end up there? Are the AI overlords fucking with people trying to fall asleep?
One is just piano, the other is keys and drums. On Bandcamp they are about 10' each, but they can be made of arbitrary length, without ever repeating themselves exactly (in principle... in practice it's likely there are exact repeats but they should be few and far between).
If you have an idea of the type of background music you need, I can make other tracks too. I'd be happy to work with you on this.
(This example is just one minute but it can be made of any length.)
Maybe the mix a little too much on the high-end, which may conflict with the voice. It should be tested?
Anyway, tell me what you think; I tried to sound a bit like the current background but with a more melodic feel and less industrial noise; but we could go in any direction -- if we know what to try.
I would never sleep with this, I would laugh too much! I love absurdist AI stories
>> So he chained her up in her room and he chained up hundreds of angry wolves in the other side of the room. [...] But he made the window and the doors big enough so that the fierce beasts could move in and out and chase her away. And they lived happily ever after.
https://www.goodreads.com/book/show/364664.The_Adventures_of...
Absolutely so. Without higher order of organizing patterns in activity, brain is just a bunch of neurons firing up - this will correlate with random phenomena in consciousness. Years ago I read that visions of going through tunnels and into the light at a distance in near-death experiences is just what visual pathways in brain will produce when activity in them fades away. 'Signal lost' experience of a (still conscious) brain. (That would be interesting to dig into this. I have no reference.)
EDIT: Actually some of the other voices are really good... I'll try that, thanks!
Definitely did not help me fall a sleep.
Also at the beginning there seemed to be one little girl (Amelia) and one witch (Sarah), and then there were now two little girls (plus the witch), and one of the little girls stood between the two little girls, and later on there appeared to be three little girls. The girl duplicating over and over sure got me hooked, kind of like watching a strange surreal painting, or reading some PKD short story.
Maybe using some sort of deepfake for voice would make this a 100x better.
I mean, she would have fallen asleep anyway, I probably could not. The voice is a little unpleasant and I concentrate too much on the nonsensical stories. But I also can't really fall asleep when the TV is running, so YMMV.
GPT3 does tend to get a bit repetitive, though, with the default temperature (0.7).
The GPT-3 generated conversations were coherent most of the time, and even interesting! However the generated speech via Google Cloud's API was monotonous and could do with a bit more intonation and excitement.
Add a few sentences very related to the child's environment. Or a word or two might be accidentally related.
It is also pretty cool though.
Reminds me a bit of Blue Jam: https://www.youtube.com/watch?v=E8VG6HUimsQ
Idea: Grammarly could make the phrasing sound a tiny bit more human.
Great idea though!
Maybe you could even position the whole thing as glued together from bits of story a fictional other-podcast left on the cutting room floor!
For some reason it gives me a really similar vibe to 88% Parentheticals, a single-episode joke podcast made for Reply All. It's just one long rambling nested parenthetical side story: https://podcasts.apple.com/us/podcast/88-parentheticals/id14...
What else could get you to sleep other than trying to make sense of things that doesn't make sense?
document.querySelector('audio').playbackRate = 2Here's that one: https://www.pastery.net/fsbjzc/
Beatles would be better. Hundreds of hungry Johns and Pauls chained in another room, ordered to make music.
https://anchor.fm/deepdreams/episodes/Episode-5-e1b6trr
With a new voice, as well!
It's kinda spooky too - listening to the thoughts of an AI!
(I know that's a stretch, but still)
https://www.stavros.io/posts/making-ai-podcast/
I posted it here but it didn't get much traction.
But I still appreciate you posting it, because it's fascinating to see how such a thing was done!
I don’t have Spotify or anchor.