Giving Stephen Hawking a voice
wired.co.uk
wired.co.uk
> "I'm trying to make a software version of Stephen's voice so that we don't have to rely on these old hardware cards," says Wood.
So back in 2010 we had someone help us extract the program ROM code from the SNES DSP-n coprocessors (used in games like Pilotwings and Mario Kart.) It turns out these chips were NEC uPD7725 DSPs. There was basically only one very terse document on how the chip worked, and no emulators for it, so I had to write one. Had a bit of help in fixing the overflow flag calculations from Cydrak.
A while later, I spoke briefly through a liaison with Sam Blackburn (who was then Stephen Hawking's assistant) back in 2011'ish. They were looking for permission to use my uPD7725 emulation code (which I said yes to, obviously.) Apparently the Speech Plus text synthesizer uses NEC uPD7720s. This is basically the same chip and ISA, but with less ROM/RAM. It's a neat little fact, but not too surprising. These DSPs are really versatile, and different programs can make them do very different things.
Reading this article, it sounds like the effort was as yet unsuccessful, though :(
(It's also important to note that the uPD7720 is probably an infinitesimal part of the overall system, so I suppose they ran into additional problems.)
Direct link to the screen-capture video: https://www.youtube.com/watch?v=mPU6mnM2i-k
TL;DR Hawking only has a single reliable, low-latency binary signal (facial muscle movements), so his interface has been a constantly-moving cursor that he can "click" when it's over the next symbol/command he wishes to select. The innovations here are in the interpretation of those selections: he now has autosuggest for text (designed specifically for him based on the corpus of his works) and shortcuts for filesystem management.
I'm looking forward to seeing when the source code is released, or when a paper is written. Just looking at the data-entry video, for instance, there are interesting parallels between the timing specifications for Hawking's Yes/No dialog and GUI design for dialogs for non-disabled users - in both cases, if there's not enough spacing in between buttons, or orderings are unpredictable, it's much easier for someone to mis-click!
The videos you linked to show this very well in my opinion:
- In the longer one with Hawking and the System he's using it to type and read Wikipedia, all the while quite some screen estate is wasted with (for him probably unusable) title bars, partially hidden desktop icons in the background and the browser partially behind his input software.
- The "data entry" video shows Notepad being opened and being partially hidden by the input UI.
That does not seem useful. I would rather use apps that automatically fit themselves to available space and predefined layouts for multiple apps or a dynamic tiling approach.
> Professor Hawking has been using his new software for several months while
> Lama and her team have been debugging and fine-tuning it. It’s almost
> finished, and when it is, Intel plans to make the system available to the
> open source community.But perhaps it's not so surprising. He's old, and past a certain age you just don't have time to relearn everything. Your time is better spent squeezing the juice out of what you have.
Huh? The quote you're replying to says he "has been using his new software for several months". Am I missing something? Is he switching back after testing? I know that he doesn't ever plan on changing the 'voice' as it has become known as his own voice but I don't see why he would test new software and then not use it. To test it you will have to learn it.
http://www.rogerebert.com/rogers-journal/finding-my-own-voic... https://www.cereproc.com/
https://www.youtube.com/watch?v=_0KUw3xr7cA
And it wouldn't be possible for Hawking. Ebert's system was put together using hundreds of hours of recordings from Skiskel & Ebert At the Movies.