Free Text-to-Speech App with natural voices
elevenlabs.io
elevenlabs.io
There is a pitfall in thinking that the most natural voice is the most important metric (i.e. many blind users still prefer espeak). Piper has a great balance of natural voice, offline convenience/privacy, and interpretability at high speeds. Still have not seen anything better as an overall audiobook solution.
Examples of Piper generally can be found at https://rhasspy.github.io/piper-samples/ , and I added my program can be found at https://github.com/C-Loftus/QuickPiperAudiobook/tree/master/...
I added this info in the readme for you and anyone else that could benefit from it.
This is what I see: https://imgur.com/screenshot-PqmyqhT
Do I need a GitHub account to see the draft releases?
They’re missing the chance to be cheap and sweep as many developers as possible into their net. OpenAIs genius is not only being first with a truly great product but being so cheap it’s the easy and obvious choice for developers.
Something better will come along.
https://www.youtube.com/watch?v=wevDlfDIG9s
It still seems a bit expensive to me, though....
Also, this bookmarklet will speak highlighted text in the browser regardless of platform:
javascript:void function(){ javascript:(function(){ var selection = window.getSelection().toString(); if (!selection) { alert("Please select some text on the page."); return; } var encodedSelection = document.createElement("div"); encodedSelection.textContent = selection; var processedContent = encodedSelection.innerHTML.replace(/\n/g, " <br></br> "); var words = processedContent.split(" "); var formattedText = ""; var speechContent = ""; for (var i = 0; i < words.length; i++) { var word = words[i]; var chunkSize = Math.floor(word.length / 3) + 1; var boldPart = "<span style='font-weight:bolder'>" + word.substring(0, chunkSize) + "</span>"; var lightPart = "<span style='font-weight:lighter'>" + word.substring(chunkSize, word.length) + "</span>"; var formattedWord = boldPart + lightPart; if (word.endsWith(".")) { formattedWord += "<span style='color:red'> *</span>"; } formattedText += formattedWord + " "; speechContent += word + " "; } var newWindow = window.open("", "_blank"); newWindow.document.write("<html><head><title>Spoken Content</title></head><body><input type='range' min='0.1' max='10' value='1' step='0.1' id='rate-slider'><p id='content' style='background-color:#EDD1B0;font-size:40;line-height:200%25;font-family:Arial'>"%20+%20formattedText%20+%20"</p></body></html>");%20var%20rateSlider%20=%20newWindow.document.getElementById("rate-slider");%20var%20utterance%20=%20new%20SpeechSynthesisUtterance(speechContent);%20rateSlider.addEventListener("input",%20function()%20{%20utterance.rate%20=%20rateSlider.value;%20window.speechSynthesis.cancel();%20window.speechSynthesis.speak(utterance);%20});%20window.speechSynthesis.speak(utterance);%20})();}();
I've played with it a bit (I made a local nonsense social media thing, reading llm content), and its surprisingly easy to get text to display in sync with the audio. It works with voices that are installed on the users system.
It's better than traditional text to speech, but I can't use it to listen to long form articles.
The app from this announcement does accept epubs. Just tested a couple and had no issues. Haven't used 11labs before, but the quality was good and didn't have any major issues with an English voice even with some spot checked fantasy names/terms or chemical names.
I have found myself finding more new books by checking which other books a specific narrator has voiced, rather than finding a book I want to listen to and hoping for a good narrator.
/rant i guess