You don't need to have it always on? This is a far cry from "$200/month," but I do not think it's $50k for "useful." Do you see it differently?
100/s*month*(.14/million) = $37
$37 for the input tokens for Deepseek V4 Flash if you miss cache all the time.A decent deal but Flash is quite dumb and you still have to pay for output tokens
I bought my spark and the models have already improved in that time (qwen3.6, speculative decoding 2x tgen, diffusion gemma 4x tgen) and I expect this to improve. Look out another 2-3 years, local is going to be very competitive.
This website is so out of touch with reality.
When a low speed of the order of one token per second is accepted, any open weights LLM can be run on an ordinary PC (with the weights read from SSDs) and the cost becomes negligible.
Such a low speed would be annoying for a chat, but I do not believe that it is "barely useful" for a coding assistant. There are plenty of tasks for which it is fine to get results some hours later or even overnight, and batching multiple tasks can complete them in about the same time as a single task.
So maybe for a hobby project this is fine, but for something you have to take to market and compete with... I think it'd be a really rough sell.
EDIT: also, just to be clear: if there was a practical path to using local AI, I'd take it in a heartbeat. I hope it gets to the point that it's better to use local than paying someone $200/mo. But right now, that $200/mo is the clear best option. I get making compromises for ideology but the compromises are too big for me right now.
Put it in front of a common-ish Python or NPM package and see what it finds, likely going to be a lot more.
Crypto bros early claims that blockchain would threaten sovereign nations' ability to collect taxes by ushering in an era of perfect anonymity to financial transactions...
Glassy-eyed consultants convincing basically everyone that introducing electronic devices into classrooms would usher in a new era of human achievement...
As a software engineer it took me a couple more decades than it should to realize that the tech industry, and especially the tech industry in CA, runs entirely on bullshit.
That very much depends on how you group your terms. Mediocre code has existed as long as programming languages, and it was something organizations trained juniors to avoid mass producing. What's revolutionary is now your boss is throwing money at tokens instead of paying for ass-seat-hours.
I'm really struggling to get a handle on your position here. You appear to be arguing the worst positions from both sides of the debate around AI at the same time. If you're claiming the future is here and it is deeply stupid then we're in violent agreement.
The future is here, but unevenly distributed. Waymo operates in a select few city, but in those cities, you can call a car, that car will have no human driver in it, and the computer will drive you to your destination. Yes it's taken a long time, but if your "evidence" is self driving cars, you might want to address your priors.
San Francisco Bay Area, CA Los Angeles, CA Metro Phoenix, AZ Miami, FL Orlando, FL Dallas, TX Houston, TX San Antonio, TX Nashville, TN
I don't know why comparing it to the GDP of Luxembourg is interesting but sure. 1.9 million people die annually in car crashes.What would you spend it on?
While I have you I'm honestly curious what it is about Waymo that attracts all the devotion? That's super not a thing where I live so I'm guessing it's got to be some kind of regional cultural thing or something?
Believe it or not, some people actually do derive a great deal of value from LLM's and it's also ok if you don't or can't.
Still feeling chippy over there I see.
Now would be a pretty good time to define "value". If folks find themselves in a position where statistically averaged word salad or time sunk combing work product for hallucinations equates value that's less an endorsement of the technology than a degradation of the term.