XTTS: Open-source Foundation TTS model
coqui.ai
coqui.ai
Followed the link. The CPML has restrictions that make it a proprietary license, decidedly not open source/free software.
Using the term "open source" in that paragraph is deceptive. Combined with the GitHub link, this leads me to believe this is just more open source cosplay.
I don't think the concept of open source translates very well to model weights. The "source" is effectively the model architecture (which is freely available) and the training scripts/data (I don't know if these are available or not). With access to those, you can reproduce the weights yourself.
IMO the situation is closer to "source is open, but if you want to use our published binaries commercially, pay us". It's not free/libre, but it's also not unreasonable. Coqui is a small company, and freely releasing the weights for commercial use would deprive them of one of the few revenue streams they have.
In such a circumstance you can still compile the source yourself. In this case, you cannot.
Also:
> Coqui is also innovating in open source model licensing.
"innovating" by making something that isn't open and falsely calling it "open source". And even that isn't "innovating" because a few others are already engaging in the same false advertising.
I think it translates pretty well. Model weights are not really that different from other non-code assets, like images or 3D models. If i can bundle open source code with such assets and offer the whole application under open source license, then such assets are open source. If the license of such assets prohibits that, they are clearly not open source.
There is question about what is 'source code' for such assets and how to apply copyleft licenses like GPL to them, but non-copyleft open source licenses like MIT/X11 do not care about that and could be easily applied on model weights or any other assets.
- Multilingual: Generates speech in 13 different languages
- Voice cloning with 3 seconds of audio
- Cross language voice cloning as well
- 24khz quality
- Blog: https://coqui.ai/blog/tts/open_xtts
- Demo: https://huggingface.co/spaces/coqui/xtts
- Model: https://huggingface.co/coqui/XTTS-v1
The speech output of this tool is as good as I have ever heard. I am looking forward to sending this to my ESP32 remote audio player.
What a world we live in!