494 karma · joined March 12, 2017
We state in our blogpost that we make an exception for Obama/Trump in order to raise public awareness. Both of them are regularly used in Machine Learning benchmarks (for example [0] [1]). Note that we don't allow users to generate from Trump/Obama's voice.
Once again, we care a lot about these issues and that's why we only allow users to copy their own voice.
[0] http://www.washington.edu/news/2017/07/11/lip-syncing-obama-... [1] https://www.youtube.com/watch?v=ohmajJTcpNk
These issues are challenging and suggestions about how you think the technology should be introduced/regulated are very welcome.
> (v) create a false identity or otherwise impersonate a third party on or through the Services;
section 3.A: https://lyrebird.ai/terms/evaluation
Unfortunately this is something that we can not enforce automatically.
We ensure that people copy their own voice by asking them to read predefined sentences (we use speech recognition to check that the sentence is indeed corresponding to the text).
To recap:
- we want to start by raising public awareness about the technology and we did demos with the voices of Trump/Obama for that,
- your digital voice is yours, people can not use it without your authorization.
And thanks, we are going to update the instructions to make them more clear.
Other interesting observation are the sentences that people generate for the first time with their digital voice...
There will be two scenarii:
- you want to use the voice of someone that has a Lyrebird account: he or she has to give you their authorization.
- you want to use the voice of someone who does not have an account. We have specific contracts for that. Say you want to copy the voice of Morgan Freeman, the contract will be between him/her, you and Lyrebird. We will also probably explore alternative ways for that.
Some other people shared their voices on twitter if you want to compare: https://twitter.com/LyrebirdAi
> It'd be nice if there was a way to tune a few parameters manually (tempo, pitch, etc).
Yes we are currently exploring ways to control the generation: volume, pitch, tempo, speed but also intonation and emotion.
What would be your use case?
1) The research of the PhD and the startup are quite complementary at the end, so we hope we can continue doing both.
2) We didn't do demo day because we raised our seed round just before YC and did not want to raise again.
Our upcoming versions should be more robust to different accents and we also plan to extend it to other languages.