13,227 karma · joined November 8, 2009
Can also just be an asymmetric bet - relatively small outlay with low chance of working, but massive upside if it does, such that it's still positive expected value.
But mostly, I think all these armchair quarterbacks should not have very strong opinions about something that they almost certainly don't know much about compared to the people who build these. The person I responded to cited a random blog that ran some numbers, and includes such gems as "In space, a broken fan or a fried motherboard is a crisis." Just... lol
But I'd also not dismiss what Tim has to say. For much longer than this stuff has been super hyped, he's reliably been one of the best sources of info on some of this stuff, especially about GPUs and how to run things locally.
Good riddance?
Very happy to have my agent totally ignore the ads trying to hijack my attention.
Seems like you all are already doing a lot of what I’ve been aiming for with mine. Are there useful ways to contribute, or have you all gotten it to a pretty good place technically, and it’s mostly a matter of spreading it at this point?
If you decide to give it a download, let me know if you have any issues with it/suggestions for improvement.
Alternatively, if someone else knows an all-in-one option that exists, I wouldn't mind retiring those crawlers...
Very much a work in progress, only federal and state so far, no municipal codes yet, and no case law yet. Big hole, I know. Also working on making the search ranking work better.
On the global market, you can get the equipment for much less, the same inverter is a fraction of the cost that the UL listed American company version costs, the panels are a good bit less. But in the US, this stuff gets marked up pretty hard.
If your main load is for summer A/C, with a mild winter, the generation and usage line up much better.
>One data point that seems relevant to me is that the previous gen qwen Qwen3.6-27b was not so different in performance from its sibling model Qwen3.6-35b-a3b. We never got a qwen3.8-35b-a3b, but if we had, would the gap have stayed the same or gotten bigger? I.e. would the quality gains by improving training coming up against a hard limitation with 35b, or not.
Yeah good question, kind of shocking that a 3b active model would perform as well as a 27b dense.
First off, I'd include Qwen flash-next and GLM 5.3 to show some of the other strong open weight models, and they predictably dominate it, but they're much larger. But, it shows up right next to DSv4 Flash 0731 on the overall index, and that's much larger. It's a great model! But then scroll down and hit Time Per Task, and you'll see that DSv4 Flash takes 3.6 seconds per task to Qwen's 21.1. That's what I meant when I said this:
>speed due to excessive thinking maybe to make up for the smaller amount of world knowledge baked in (qwen 27b's main issue iirc), etc - they're tuned for different things.
It can make up for its shortcomings by iterating a lot longer, and using way more thinking tokens. And that's a great trade if you don't have the vram to run the bigger models, but speed is pretty important for getting things done... And that's why DSv4Flash is great, too, despite being much larger, and scoring similarly on the intelligence index.