Silent speech with ultrasound
alephneuro.com
alephneuro.com
With wearable devices becoming more common - I am anticipating a wave of "sensors" that can be as simple as small band-aid patches that wirelessly send data to your smart device. Those sensors could also open up human-coputer-interface innovations like these.
---
In similar space ...
There was a post about a thought-to-text project from MIT no less --
"AlterEgo"
10 months ago: https://news.ycombinator.com/item?id=45174125 8 years ago: https://news.ycombinator.com/item?id=16780357
When I saw the demo posted last year it left me with an uneasy feeling -- gut feel said it was more marketing than a real working technology demo. Nothing seems to have come out of that lab since then -- strengthemning my suspicions.
[1]: https://scispace.com/pdf/sub-auditory-speech-recognition-bas...
Not sure I'd want to put an adhesive patch on my neck every morning so I can silently talk to an LLM in the cubicle farm. I hope this is not our future.
Very cool tech though and surprisingly good results for so little training.
I think time might be better spent improving a lip reading model (no adhesive required), assuming we're unable to read brainwaves directly.
The idea is that you won’t need cubicles anymore. Your hands will always be free because you won’t need a keyboard or mouse. You’ll be able to control the digital world using only your tongue.
The future is bright: you’ll talk to your Jarvis while driving through the Norwegian fjords.
It could also be the ultimate, always-on remote control for everything around, with a near-zero error rate.
It could help not only people with vocal cord problems, but also essentially anyone who is paralyzed from the neck down.
I am researching in this field right now, MY ideas are if there is minimum amount of muscle movement will be the best .
Model already generalizes on out-of-distribution people, so you actually don't need to train a model for each person
See this hn thread about it: https://news.ycombinator.com/item?id=46816228
I have never liked talking aloud to Siri or similar, but I could see using "voice" as an interface for so much more if I could speak silently to my device.
The funny part is that the tongue itself could act as a cursor or perform gestures, which sets this approach apart from other silent-speech methods.
It ended up altering my handwriting even after I stopped using it.
That’s a remarkably small number compared with other methods. English could easily be optimized to achieve a sub-5% WER with a very small model.
After the session, I tried to use Wispr Flow but forgot that I was still speaking silently, so it didn’t work.
When I finally said it out loud, it felt like such a hassle that I essentially sold myself on the idea of silent speech.
In the office, a non-contact video solution (lip reading) is likely to be far more popular, but a lot depends on which is more accurate.