Is the data going to be freely available as well? It's a little unclear whether they intend to make it separately available or not.
The relevant blurb:
Your Contributions and Release of Rights
By submitting your recordings, you waive all copyrights and
related rights that you may have in them, and you agree to
release the recordings to the public under CC-0. This means
that you agree to waive all rights to the recordings
worldwide under copyright and database law, including moral
and publicity rights and all related and neighboring rights.
[1] https://voice.mozilla.org/termsI'm wondering if the format will be easily translatable to the kinds of models that software like CMUSphinx and Julius use
If you poke around github and the Kaldi lists a bit more you can see that they are experimenting with and probably planning to use Kaldi.
I wonder what they plan to do for provisioning. It is one thing to collect data and train models, but quite another to make the service available over the web in an unlimited capacity. And we are not yet to the point where you can reasonably expect to run a high quality open-vocabulary STT system in your browser. The search network is typically in the GBs range.