I'm sure mainframe time-share providers in the '60s and '70s were salivating at the possibility of computers mediating most business tasks, too, completely unaware of the microcomputer revolution that was about to happen.
I'm sure mainframe time-share providers in the '60s and '70s were salivating at the possibility of computers mediating most business tasks, too, completely unaware of the microcomputer revolution that was about to happen.
I don't know how much experimentation you're doing with local AI, but that day may be sooner than you think. The ecosystem is evolving extremely rapidly.
Looking forward to the day this becomes very very easy for the likes of me.
In fact, we're already further along than that in terms of local AI. I'm currently able to get usable results at 8-10 tokens/sec using open-weight models on my laptop's integrated GPU, running on battery power. A $4,000 DGX Spark (less than what an IBM PC cost at launch in inflation-adjusted dollars) can get 3-5 times the inferencing performance with models 3-5x larger.