https://github.com/facebookresearch/wav2letter/blob/master/r...
My models are here: https://talonvoice.com/research/ I haven’t yet posted the model I’ve been working on most recently. I’m 120 epochs into a large size model trained on all of my datasets. I also have another 1000-1500h of audio I haven’t finished prepping to train on.
Here is a web demo. It’s currently running a slightly older checkpoint of my WIP large model, and the deepspeech LM: https://web2letter-west-1.talonvoice.com/
Do you have any experience with online decoding in wav2letter ? Is there something like a Websocket API available somewhere ?
the english is from the tedlium recipe with WER of 7%.
room for improvement, but for our original purpose it was sufficient.