0 - https://www.williamangel.net/blog/2026/05/17/offline-llm-ene...
If inference cost comes down (as it has been for the last few years) you’ll be able to run today’s SOTA in your laptop by the end of the year.
I want local AI to be a thing but the hardware isn’t here yet, because the only options are a Mac Studio or DGX machines strapped together. RAM prices needs to crash before local AI has a chance at actually competing.
As soon as I can buy hardware for less than 5k that runs an opus 4.6+/5.5 model locally I will do it instantly
If Claude hosted on AWS bedrock is not considered trustworthy, I have some bad news for you.
How is that going to work on Bedrock, when they don't even manage the infrastructure?
And, who wants to screw around with harnesses or define agent orchestration when Claude/codex are good at this and getting better every month.