HNHacker News
TopNewBestAskShowJobs

_josh_meyer_

438 karma · joined January 3, 2022

building
submissionscomments
_josh_meyer_··on Show HN: Voice Clones for Creators
same company, yes, but major updates since our demo:

1) Quality: the underlying model is much, much better, and so is the quality of the Voice Clone. 2) Productization: previously we just had a stand-alone demo, now we've launched the product (user accounts, multiple voices, etc.)

_josh_meyer_··on Show HN: Voice Clones for Creators
yours is a solid point -- would be nice to have that "wow factor" right off-the-bat
_josh_meyer_··on Show HN: Voice Clones for Creators
still getting the error? I can't replicate on MacOS Monterey
_josh_meyer_··on Show HN: Voice Clones for Creators
delicious documentation for your reading pleasure:)

https://tts.readthedocs.io

_josh_meyer_··on Show HN: Voice Clones for Creators
thanks for linking -- the core project is indeed open source! the founding team all worked at Mozilla on TTS and DeepSpeech, then we spun out to create Coqui <3
_josh_meyer_··on Show HN: Voice Clones for Creators
Product Hunt --> https://www.producthunt.com/posts/coqui
_josh_meyer_··on Show HN: Voice Clones for Creators
more fine-tuned control over enunciation/emotion and the like is very much in the works... stay tuned :)
_josh_meyer_··on Show HN: Voice Clones for Creators
Thanks, Neil :D The current model definitely does an "implicit" accent conversion to American English, but in the near future we'll have something that keeps accents on separate sides of the pond, so to speak :)
_josh_meyer_··on Show HN: Voice Clones for Creators
the voices are demoed for [videogames](https://www.youtube.com/watch?v=x8tEdwll_CY&t=4s), [dubbing](https://www.youtube.com/watch?v=TBjdUY3_ccQ&t=1s), and [post-production](https://www.youtube.com/watch?v=ykoBFrb9itY)... but I see your point - thanks!
_josh_meyer_··on Show HN: Voice Clones for Creators
thanks for flagging!
_josh_meyer_··on Show HN: Voice Clones for Creators
why don't you try :D ? Make a clone where you're talking like Hal... see how it does ;)
_josh_meyer_··on Show HN: Voice Clones for Creators
currently it's only user-facing, but still early days:)
_josh_meyer_··on Show HN: Voice Clones for Creators
of course! we've got an open ethics discussion going here: https://github.com/coqui-ai/TTS/discussions/1036

and our chatrooms are probably the best/fastest way to get ahold of us: gitter.im/coqui-ai/TTS

_josh_meyer_··on Show HN: Voice Clones for Creators
video game demo --> https://www.youtube.com/watch?v=x8tEdwll_CY
_josh_meyer_··on Show HN: Voice Clones for Creators
walk-though video --> https://www.youtube.com/watch?v=ri1U-bJ-6vc&t=4s
_josh_meyer_··on Show HN: Voice Clones for Creators
thanks for flagging! what OS are you on? iOS seems working OK atm
_josh_meyer_··on Show HN: Voice Clones for Creators
For this zero-shot approach, there's actually no "fine-tuning", and from what we can tell ~30 seconds is optimal (the text isn't actually even used... you can say anything!)

For longer outputs and legit fine-tuning, that could become an offering :D

_josh_meyer_··on Show HN: Voice Clones for Creators
Voice Clones for:

-> Video Games https://coqui.ai/video-games -> Dubbing https://coqui.ai/dubbing -> Post-production https://coqui.ai/post-production

Free access

_josh_meyer_··on Show HN: Clone your voice and speak a foreign language
very much intentional.

Background music makes misuse/abuse less likely (both intentional and unintentional)

Read more here about in our open discussion: https://github.com/coqui-ai/TTS/discussions/1036

_josh_meyer_··on Show HN: Clone your voice and speak a foreign language
Demo: https://coqui.ai Code: https://github.com/coqui-ai/tts Blogpost: https://coqui.ai/blog/tts/yourtts-zero-shot-text-synthesis-l... Paper: https://arxiv.org/abs/2112.02418
← PreviousPage 2 of 2