Send Data with Sound
github.com
github.com
- Shall we switch to audio data for more efficient communication?
- Yes. [MODEM NOISES START]
That said, the "gibberlink" demo is definitely much slower than even a 28.8k modem (that's kilobit). It sounds cool because we can't understand it and it seems kinda fast, but this is a terribly inefficient way for machines to communicate. It's hard to say how fast they're exchanging data from just listening, but it can't be much more than ~100 bits/sec if I had to guess.
Even in the audible range you could absolutely go hundreds of times faster, but it's much easier to train an LLM that has some audio input capabilities if you keep this low rate and likely very distinct symbols, rather than implementing a proper modem.
But why even have to use a modem though? Limiting communication to audio-only is a severe restriction. When AIs are going to "call" other AIs, they will use APIs… not ancient phone lines.
The original plan was to develop essential "audio QR codes" that would allow short codes to be transmitted that could be parsed by certain apps and used to drive different interactions.
Does some device listen for apps nearby? Do I need to walk up and press a button?
If it had all gone off well, the eventual plan was to have it be used on a live show where users could also interact. We had some prototypes ready with a native app - but then the Brexit referendum happened - and our company had a couple of clients pull out of upcoming projects - and the company got shuttered.
Ironic that the author overlaps so much with that field, without noticing that they chose the same name as probably the most used amateur radio programmer in the world.
If you're interested, the state of the art is VARA. It's closed source though, so NinoTNC may be a more interesting choice.
I'm not a lawyer, nor is my ham license even in the US, but perhaps "you can decode it by using our software" satisfies the legal requirements?
It's not, to my knowledge, deliberately obscured. That would be a legal no no, I think.
But yes, people have fought over VARA's state here.
> standard FSK protocols such as Bell103, Bell202, RTTY, TTY/TDD, NOAA SAME, and Caller-ID
E.g.,
printf 'Hello, world\n' | minimodem --tx 440
minimodem --rx 440
(you can choose any freq.) results in a lot of, ### CARRIER 440 @ 800.0 Hz ###
�
### NOCARRIER ndata=1 confidence=1.507 ampl=0.060 bps=439.96 (0.0% slow) ###
### CARRIER 440 @ 800.0 Hz ###
�
### NOCARRIER ndata=1 confidence=1.858 ampl=0.053 bps=439.96 (0.0% slow) ###
### CARRIER 440 @ 800.0 Hz ###
�
### NOCARRIER ndata=1 confidence=1.832 ampl=0.063 bps=439.96 (0.0% slow) ###
and even when it does hit, ### CARRIER 440 @ 800.0 Hz ###
Helln, world�
### NOCARRIER ndata=14 confidence=2.939 ampl=1.167 bps=438.67 (0.3% slow) ###
If I try something like the example where he cats a man page: ### CARRIER 1200 @ 1200.0 Hz ###
��-O���܇����������������������=����~`���|�����������������������������_��������=����??�����?�����oﯰ������������������|���������������������߿��������������������������������������~�����`�|�w������������-Ӱ��>��み����>�����
… I'm in a quiet room.Anyone?
10 symbols per second