Also piping in and processing the data from my mobile would be cool, but I wouldn't want to invade other people's privacy if I'm in public.
Also piping in and processing the data from my mobile would be cool, but I wouldn't want to invade other people's privacy if I'm in public.
$5/month's worth of "cloud" is going to work out to be less actual raw CPU resources than a low end raspberry pi running full time in-house
One second of Google cloud TPU has roughly the same number of floating point operations then 4 hours of raspberry pi 4B time.
So 3 minutes of cloud TPU time already covers your whole month of raspberry pi usage. Pretty sure it costs them less than 5$ as well, since they have the hardware anyways.
If Google (or whomever) needs to run voice models, they take your query and all the other queries that arrive in the same millisecond, smoosh them all together and shove the batch into a TPU and run it. You don't have any TPUs and you also don't have any traffic you can use to amortize the cost of your infrequent queries.
The idea that you could run these kinds of ML inference tasks is economically fanciful. You would need a huge investment in hardware and the opex would be ridiculous.
Google, Apple, Amazon and even Sonos are all releasing voice assistants that work locally on their relatively low powered speakers.
Apple seems to be ahead with what is local, while Google seems to be the smartest. (Sonos doesn’t have a cloud, but it’s not ‘general purpose’ afaik).
Sure you can’t amortize them across a bunch of TPUs BUT instead they can ship custom hardware. A tpu needs to be big and support parallel streams. A home server may only need to ever serve one stream. There are arduino style devices that can perform basic tensor flow audio models in real time now. And obviously most phones can perform this locally now, so depending on opinion that may be considered affordable.
Searching all of a downloaded copy of Wikipedia wouldn't be that computationally expensive either if the assistant has hot words it picks up to look up.
Could just load up a pcie card full of them if necessary. A local home AI would be such a boon to people, not just the average person but the elderly as well, combined with a refined GPT etc it could conversationally respond to requests rather than most assistants' current request->response "I am a robot" scheme.
>Your son called when you were asleep to ask if you wanted to get coffee today, shall I call him back for you or put you through to him? >X, you've fallen! Please let me know you're okay or I will call emergency services for you
It's sad that we have the technology to do this already but haven't.
I suspect OP was clear enough.
But there exists https://mycroft.ai/