I’m more curious on the size. If it’s smaller than or equal size to GLM 5.3, this would be a crazy good model. If it’s closer to deepseek pro, it would be a good model. If it’s near Kimi K3, I think it’s competitive but nothing particularly differentiating.
GLM series has made it very practical to self host. If the new update for Deepseek flash holds up, I think it would be silly for some companies to not self host.
It’s wild how broken search has become, LLM made it worse but it was already downhill prior to that. Even Reddit forces you login to search within a subreddit and displays the annoying “Best place on internet” overlay after staying on the site for 2 min. I just use AnythingLLM with DuckDuckGo + local LLM model to do searches now.
Overkill is the point. Same thing for iPads and iPhones. It gives a luxury feeling. That and Chinese manufacturing have made milling a very cost effective manufacturing method.
iPhones are iPads have a milled CNC “unibody” as well… Being able to dictate where to have more thickness helps a ton on the luxury feel. Most of the chips are recycled anyway into new billets.
Thanks for the update! I wonder if a Blackwell GPU would be noticeably faster. Which vendor did you end up using? I want to get a gigabyte one but thats be OOS for months.
I think the hardest part is how to define “waste”. Are you “wasting” your time if you go to a school event for your kids during the weekday? Is it a waste of time walking your kids to the bus? Is it a waste of time reading to my kids before they sleep? I argue it’s extremely hard to draw that line and a personal question for sure but being “efficient” is hard.
The issue with “AI” is how toxic it feels. Feels grimy like social media. Just like social media, I haven’t introduced any of it to my young kids nor do I plan to until they’re teens. At most I give them access to local AI through home assistant. I have a coworker who have kids of similar age and had to get rid of all his echos when Alexa started feeding his son new Lego sets as “Christmas gift ideas” to tell his parents.
Because distribution prices are spread throughout the users. This is the case for policies that allow the rich to be completely off grid which is not the case here in the US (most cap off around 1500W).
No way. I cook everyday and induction is vastly superior than gas. The only time fire is vastly superior is when I use a wok which I have a stove I use outdoors. Most induction stoves have temperature maintained too. So it can keep your oil at a specific temperature let it be frying or confit or roasting spices.
I’ve been using induction since I have had kids and I enjoy it so much more than fire. That being said, I rent so I can only use 120V ones and the options are so limited here. Most of the small 120V ones make an obnoxious hum.
I work in embedded space. Just because it’s in hardware doesn’t mean you can’t “protect” it. Most modern software (regardless if it’s hardware or not) can be cryptophically signed.
I don’t get why this is an issue? You can run Claude/OpenAI SOTA models through Amazon bedrock. These weights have to live somewhere to run on Bedrock.
SOTA American models are not. SOTA Chinese models are. From a physics aspect, closed source models cannot be too far from open source ones in terms of size. There’s only so much you can squeeze out a B100 style cluster even with fancy Dflash style diffusion model for the speculative model.
Honestly, this is starting to make more and more sense. SOTA models are starting to converge to certain architecture and capabilities. I wouldn’t be surprised we end up with a base model ASIC + “fine tune” card where it’s a physical LoRA style adapter.
For personal stuff, I use it with AnythingLLM. It replaced any Google search for me. For coding, I run opencode though I have been debating switching to Pi. I would argue it’s at Sonnet 3 level.
I meant Qwen3.6. Unsloth supposedly has early preview of the model and the VRAM requirement is the same so most people expect similar model size and type.
This reminds me of Three body problem and how the scientist discovered the high strength wire was through quick physical experiments and use them as input to an AI model to determine if it works.