Coqui TTS: a deep learning toolkit for Text-to-Speech
github.com
github.com
I've been using Google TTS for generating audio for my reading list, this would be good time to build a simple api+worker wrapper around this and integrate into my app.
Mimic is another one
Anyone who has ever fallen asleep anywhere in Puerto Rico will probably be quite familiar.
I used Coqui TTS a few months ago to roll my own speech controlled desktop in an hour or so, very cool stuff.
And that makes you a danger on the road when you're driving and you haven't been able to sleep for an entire freaking week, or maybe even a month or more.
Who would have thought so much damn noise could come from such a tiny frog!
[1] https://github.com/mozilla/TTS [2] https://github.com/coqui-ai/TTS/tree/e9e07844b77a43fb0864354...
FYI the format that they expect in metadata.csv has changed over time, it used to be "filename|transcribed text" and now it expects "filename|speaker name|transcribed text" but that's not reflected in the notebook.