Going along with this: What are the latest and greatest open source speech-to-text models and/or tools out there?
Would love to hear from experienced practitioners and a bit of detail on the experience.
Thanks HN community!
Would love to hear from experienced practitioners and a bit of detail on the experience.
Thanks HN community!
Mozilla Deep Speech is an open source speech recognition engine, based upon Baidu's Deep Speech research paper[2].
Unsurprisingly, Deep Speech requires a corpus such as... Common Voice.
It also looks like Baidu are now developing their Deep Speech as open source? https://github.com/PaddlePaddle/DeepSpeech