Nice. The voice sounded completely real. Are you using AI on the back-end or real humans? If real humans how long is the turnaround time?
Multiple orders of magnitude.
TTS and VC are going to do to voice over and narration what Getty and Flikr did for photography.
We're not quite there, but we're not far off. Give it two more years.
(I work pretty closely on this domain. I built https://vo.codes, which sounds like shit, but it's adjacent to cutting edge techniques that sound practically real.)