It becomes another abstraction, really. As long as we can use it for something useful, it's still valuable.
1,209 karma · joined November 6, 2013
It becomes another abstraction, really. As long as we can use it for something useful, it's still valuable.
Yes. There are multiple public documents showing this, ex:
> The FAA has determined it is not appropriate to mandate the use of CRSs in aircraft now. We remain concerned that if we require children under 2 years old to be in an approved restraint system (which requires a passenger seat), affected operators might find it necessary to charge a fare for transporting these children. (Currently most, if not all, operators do not charge a fare for children under 2 years old who are held in an adult's lap.) In turn, for economic reasons some adults might decide to drive in automobiles to their destinations rather than fly. The FAA is concerned because automobile injury and fatality rates are higher than aircraft injury and fatality rates. As a result, there would be a net increase in transportation injuries and fatalities as families opt, for economic reasons, to drive rather than fly to their destinations.
https://www.federalregister.gov/documents/2005/08/26/05-1678...
> * Subsidize the 2nd seat for parents with children from tax dollars
That's not something the FAA or DOT can do.
> * Criminalize many intentional driving violations
That's not something the FAA or DOT can do.
> * Build high speed rail
That's not something the FAA or DOT (unilaterally) can do.
> 2) It generates structured output natively - guaranteed to be correct
It's not guaranteed to be correct: it's guaranteed to be _formatted in a particular way_. You can get the same thing with grammars on any LLM.
Jev and Jev-like models have other advantages, but I feel like people forget grammars exist for LLMs.
I don’t see any reason to think the same thing that happens with all tech won’t happen here.
I understand this, but I just can't bring myself to care. I've been doing professional software work for almost two decades. These sorts of improvements/time savers are great without AI. With AI? Whatever. It's fine.
When the underlying lib has an issue, it'll be quicker to debug with the whole thing in context.
This is at a well-known tech company operating at massive scale (and resulting complexity).
L1/L2 tech support is completely dead within a couple years. The delay is only around how long it takes people to realize.
Not a reasonable assumption for a variety of reasons.
> 2. Deepseek is served at a profit and not a loss
Not a reasonable assumption either.
> Why do you need to know the architecture? Just compare Deepseek V4's performance with GPT 4 and treat internals as a blackbox.
Because the internals are what actually matter and what drives inference cost.
It would be entirely reasonable to expect that GPT-5.5 has some sort of optimizations or changes to the architecture to make it easier to train, or to make runtime ablation easier, or to better handle large batches, or whatever.
Those changes, particularly if they are non-public, can easily result in worse inference performance than a comparably sized model without those changes.
> It is borderline conspiratorial to believe it this way.
It's not any sort of conspiracy. It's how land-grab tech companies have always worked. To presume otherwise is silly.
It would be more surprising if the surrounding architecture hasn't significantly diverged. If it _hasn't_ significantly diverged, then given the performance difference it would imply that the frontier models have significantly greater param counts, which would result in a higher cost.
That blog post is not very compelling either. Without knowing details of the architecture, comparing the various frontier models to open models doesn’t make sense.
I guess (heh) it depends on your definition of 'educated guessing'? Looking at the problem, considering a solution, discarding it, trying another and testing, iteratively, is how most people would approach any tricky problem.
Brute force is substantially different. It would be saying that, other than maybe setting some basic bounds and heuristics, I'm going to try literally everything and test each. That's not at all what the LLM did here.
Sounds like Europe is behind because Europeans are working less and taking more vacations. You just point to poor leadership as the cause.
https://www.dni.gov/files/ODNI/documents/assessments/ODNI-Un...
https://www.dni.gov/index.php/newsroom/congressional-testimo...