https://venturebeat.com/2021/04/12/mozilla-winds-down-deepsp...
https://blog.mozilla.org/en/mozilla/mozilla-partners-with-nv...
$1.5mil for shutting down open source initiative, almost half of CEO salary right there.
https://venturebeat.com/2021/04/12/mozilla-winds-down-deepsp...
https://blog.mozilla.org/en/mozilla/mozilla-partners-with-nv...
$1.5mil for shutting down open source initiative, almost half of CEO salary right there.
Especially for the videos with Close Caption....
As simple as extracting the Audio and CC text?
(See https://support.google.com/youtube/answer/2797468 and the part about status.license here: https://developers.google.com/youtube/v3/docs/videos)
Anyone can use the Common Voice data within the terms of the license and NVIDIA contributing towards the continued gathering of data (that will continue to be made publicly available) won't change that.
It's a huge shame that Mozilla didn't continue the DeepSpeech project but Coqui is taking on the mantle there and there are plenty of others working on open source solutions too, all whilst the existence of CV will make a big difference to research, in the academic, commercial and open source spheres.
If that was true that would be a profoundly bad purchase for NVidia since the data is already freely licensed and available for anyone to use at no cost.
This is like saying that Epic "bought" Blender when they gave it a development grant, or that Google contributing patches to upstream Linux means they own it now. Mozilla didn't give NVidia any kind of special license, when NVidia contributes data to Common Voice they're doing so under Common Voice's license, not their own.
We want to encourage more companies to treat software and training data as a public commons that is collectively maintained, this is a good thing.
https://techreport.com/news/14707/ubisoft-comments-on-assass...
https://techreport.com/review/21404/crysis-2-tessellation-to...
https://arstechnica.com/gaming/2015/05/amd-says-nvidias-game...
Here it appears they purchased this https://venturebeat.com/2021/04/12/mozilla-winds-down-deepsp...
And the assumption the shutting down Deep Speech was specifically for NVidia's benefit seems like a fairly large leap to me, given that Deep Speech is already mature, still being developed under Coqui.ai, and surrounded by a wide diversity of other deep learning projects that also aren't controlled by NVidia.
Decreasing barriers of entry for those models and providing raw data is probably the right thing for Mozilla to be focusing on right now. Any team can build a language model, only companies like Mozilla can coordinate mass data collection for those models.
I think real-time factors smaller than 1 are faster than real-time (not slower) and use less than 100% of a resource's computational power to keep up.
> I think real-time factors smaller than 1 are faster than real-time (not slower) and use less than 100% of a resource's computational power to keep up.
Sure, but who has the necessary GPUs installed? And on CPUs it will apparently take longer to generate speech than the duration of that speech. Unusable for many UIs and it will also drain the batteries of any portable device.
So it is the contrary
It's significantly closer to "nonfree" on the free-nonfree spectrum than it should be, and is another example of the difference between the guiding philosophies behind "free software" and "open source"