- https://github.com/gooofy/zamia-speech#asr-models
- https://github.com/mpuels/docker-py-kaldi-asr-and-model
in regards of speech recognition except the fact that its easier to use?
- https://github.com/gooofy/zamia-speech#asr-models
- https://github.com/mpuels/docker-py-kaldi-asr-and-model
in regards of speech recognition except the fact that its easier to use?
the other one is an example for packaging kaldi in a docker container.
in the past to provide Speech training data, but they were not really interested. I thought it would be great to improve STT having a real good and HUGE set of german audiobooks based on Text, that is publicly available... unfortunately i had no success trying to script something for this purpose (mainly lack of time).
It basically describes the thing you mentioned - matching freely available audio books with the source text and using some tools to preprocess the data suitable for ASR training (alignment, splitting).