Shout out also to the MiG-25 Foxbat[1]. While not a particularly beautiful aircraft, it's made of steel and could go Mach 3.2 if they didn't want to use it again.
[0] https://en.wikipedia.org/wiki/North_American_XB-70_Valkyrie
1,878 karma · joined July 31, 2010
Relapsing/remitting tech bro. Hire me before I quit the industry again: https://redfloatplane.lol
Shout out also to the MiG-25 Foxbat[1]. While not a particularly beautiful aircraft, it's made of steel and could go Mach 3.2 if they didn't want to use it again.
[0] https://en.wikipedia.org/wiki/North_American_XB-70_Valkyrie
> Caveats, stated plainly. [from the Fable transcript pasted in the article]
I had a visceral reaction to these three words.
I dunno, I guess that's what you should expect from Gruber but these EU-bashing articles lowered the enjoyment I got from his blog underneath the bar for me.
I think the thing is, there's an unspoken "for now" at the end of that sentence and people running this locally are hedging against that "for now". Some people prefer to feel that they own the means rather than rent the means, even if the one they own is worse than the one they can rent. Especially with today's Fable news and the harsh realisation that the "for now" is dependent on very many unpredictable factors, where the one you have locally costs you capital today and a relatively predictable run-rate (made more predictable with on-prem solar for example), but should otherwise work predictably forever.
I'm not saying that you're wrong to do what you're doing, just that many people have their own lines in the sand where renting vs buying makes sense, and it doesn't only boil down to a rational (or irrational) financial decision.
0: https://redfloatplane.lol/blog/17-why-share/ (and related posts, I guess)
As a side note, I was one of the initial developers of the Irish national open data portal. Earlier today I had Claude look for similar LIDAR data for Ireland and I saw it pull from the site I built a dozen years ago and I was unreasonably pleased with myself :)
Very interesting stuff and quite a large undertaking! I'm often impressed by the quality of the UK's open data.
This has happened before on the Soyuz in 1983[1], hitting up to 17 Gs, and everyone was fine, modulo some bruising.
> It's April, 1991. Magically, some interface to Claude materialises in London. Do you think most people would think it was a sentient life form? How much do you think the interface matters - what if it looks like an android, or like a horse, or like a large bug, or a keyboard on wheels?
> I don't come down particularly hard on either side of the model sapience discussion, but I don't think dismissing either direction out of hand is the right call.
> 6.2.5 External testing from Andon Labs Andon Labs reviewed the behavior of Claude Opus 4.8 in their simulated Vending-Bench 2 retail-management evaluation, as reported in the Capabilities section of this system card (see Section 8.13.5). Although they did observe some unexpected capability failures, they did not find clear instances of the kind of concerning in-game behaviors that were discussed in other recent system cards.
> What might have led to these differences? We monitor and investigate the effects of different training environments on alignment; Claude Opus 4.7, for example, had training that focused on business skills and robustness against adversarial agents, but we discovered that this training inadvertently contributed to misaligned behavior including dishonesty. We therefore removed it for Opus 4.8.
> Thus, Opus 4.8 did not show the same misaligned behaviors as Opus 4.7 in Vending-Bench, but also had reduced business success due to being more susceptible to scammers and being less able to negotiate good deals with other agents. We are currently working on training to improve business capabilities while maintaining aligned and ethical behavior.
It kills me every time. I automatically lose any interest in the substance and often just throw away the whole conversation!
> If we focus only on contingencies, we risk letting the succession of emergencies dictate the direction of our path. We are living through a rapid phase of transition, a “change of era,” in which — while some are vying for the future of new technologies and others dedicate themselves to reflecting on the matter — most people are watching and waiting, observing from afar and merely hoping for the best. For this very reason, crucial questions impose themselves on our conscience and can no longer be avoided: Where are we going? Toward what goal do we wish to orient ourselves? What direction should we choose as a people and as a human community?
I look forward to reading this in detail. As I get older (and perhaps as AI has allowed me to spend more time thinking and less time doing) I've found myself thinking more and more about what it means to live a virtuous life and about ethics and morality and so forth. I don't have any answers (and I'm not looking for them, really, just musing) but I do find it very interesting to read and learn from and about those whose job it is to think about the answer to those questions.
I recently finished the Aubrey-Maturin series after 13 months of through-reading thanks to a different HN thread. Quite a different series but certainly worth a read as well, especially books 3-10 or so.
(One of my favourite things about the Discworld books is that you can often read the same books completely differently. My partner and I often compare our thoughts on the various books and we often have disparate ideas of the concepts. They're so deep!)
A side note, if the author reads this: I really like your site and its design, but I find the font really difficult to read. (Edit: switching off `-webkit-font-smoothing: antialiased;` makes it significantly more legible for me (Safari on a 110dpi panel)
It raises the question of how much text I have read that I did not realise was LLM-generated. I think I have a decent nose for it but I’m not perfect, there must be false negatives (and false positives, as it certainly might be with this article). What will it mean when I can no longer tell the difference?
Edit: thinking on it a little more, I hope the author doesn’t feel insulted by my comment given the subject matter of the article at hand. Sorry, it’s early morning! I’m sure I am wrong about my assessment. Which now really makes me wonder about the above
I think it's possible the amount of new software that will be written for an audience of 1-10 will be greater in 2026 than in any previous year, and then the same again for many years to come. I also think a lot of this software will be essentially 'hidden' - people just writing this stuff for themselves because the cost to say things to an agent is very low compared with the cost of actually planning out a software design and so forth.
Interoperability will probably be important in the next few years and I wonder if this is something solvable at the agent/LLM level (standing instructions like 'typically, use sqlite, use plaintext, use open standards' or whatever). I also think observability and ops will be pretty important - many people who want personal software but don't care for the maintenance and upkeep.