Software Automatic Mouth – Tiny Speech Synthesizer
github.com
github.com
https://github.com/boourns/SAM
I then ported it to Mutable Instruments Braids, a eurorack module:
If you want to hear the output you can directly go to
https://simulationcorner.net/index.php?page=sam
The synthesizer was even able to sing: https://simulationcorner.net/SAM/sing.wav
(EDIT: author refers to it as abandonware.)
A few of my first assembly language hacks was to get it working when in 80 column mode and getting it to work in ProDOS.
The original version worked by redirecting text output to it and then you had to call another routine to redirect text output back to the monitor. But it would always revert the screen back to 40 column mode.
The second hack involved copying it to a ProDOS disk and changing the output to use the ProDOS routine for text output.
Is it true the Apple // version came with a hardware card?
It worked both with my original Apple //e and later my Apple //e card for my LCII.
In some ways I like this resurrected SAM more than things like Rsynth. RECITER seemed like magic at the time, though now with a linguistics degree, I prefer raw SAY with phonemes.
I love this era of speech synthesis. Robots were white and chrome and would one day be our cheerful servants. That they talked/sounded different from us only endeared them to us — reminded us that they were after all machines.
Siri and her sisters, like Pinocchio, want to be real humans. I would be surprised if in 40 years anyone will be nostalgic for them.
I had collected a bunch of General Instrument’s SP0256 chips that I sold (along with the rest of my electronics hobby stuff) maybe 15 years ago. I think there was a Basic Stamp or two, a Handyboard, and maybe a Pic chip or two.
And thanks for reminding me of the awesome Handyboard and it's "Interactive C" language.
As in, orders of magnitude bigger. SAM needs only a few tens of kB and produces recognisable but very robotic speech, whereas the ones you've linked appear to be hundreds of MB and probably produce something closer to an actual human. I wonder if it's possible to have very convincing procedurally-generated human-like speech in a size larger than SAM, but less than those ML methods; e.g. hundreds of kB, or few MB.
Here's a formant synth on the same order of magnitude as SAM, but size-optimised by the demoscene: https://www.pouet.net/prod.php?which=50530
https://ataripodcast.libsyn.com/antic-interview-385-software...
it converts speech to a waveform which you can modulate as you wish