HNHacker News
TopNewBestAskShowJobs

SilverBirch

7,327 karma · joined May 11, 2022

submissionscomments
SilverBirch··on You Know Who Hates AI? Insurance Claims Adjusters
So I have three thoughts about this: The first thought is this is likely mostly survivorship bias. Yes, the claims that get to the humans are the ones that went wrong, and the mode of failure for LLMs is less that they get within 99% all of the time, it's more like 1% of the time they do something insanely dumb. At which point your customers are pissed off. That's kind of obvious, and it sucks for the humans who have to clear up the situation. But it's still a win for the vast majority of cases.

Secondly, does anyone actually know how much of your insurance premium is going on the call centre staff? most people don't claim, those that do mostly have simple claims. The actual cost for insurance companies is actually paying out claims, so even if you automated all of this it likely doesn't change your economics. Historically "New" Insurance companies only really succeed by mis-pricing risk and gaining market share that way - they often then fail when that risk materializes. Lemonade sounds a lot like that.

Finally, they haven't engaged with the tidal wave that's coming. Sure, AI agents handling claims is happening now. Just you wait. In a couple of years time it'll be AI agents making the claims for you and suddenly there'll be bots filing claims with infinite patience and a direct mandate to try and get as big a payout as possible. That is when it's going to get really hairy.

SilverBirch··on How accurate have Ed Zitron's AI skeptic predictions been?
This is kind of a "I'm not here to tell you about Jesus he either lives in your heart or he doesn't".

I am an engineer, I have about 15 years of experience. I have been using AI in my job since mid 2025. Over that time it has gone from being an interesting toy that could kind of help but would often hinder, to being an absolutely explosively powerful tool. Just from personal experience it is the thing that has improved the most of any of my tools in career. And over that time my spend on AI has sky rocketed.

You don't have to believe me, it's obviously just anecdotal, but for anyone in the same position as me (And there really are lots of us), to claim model capability peaked in 2024 is just staggeringly dumb. It'd be like claiming electric cars peaked in 2008. I don't know how further to convey this to you.

It may very well be the case that the financial side is a bubble that horribly bursts. But the technology is real and the claims Ed has made are just wrong.

SilverBirch··on How accurate have Ed Zitron's AI skeptic predictions been?
For someone who falls in the middle of this a really good read is Quoth the Raven[1]. He essentially makes the argument that proposition 1 is probably true, but that before proposition 2 happens there'll be a massive bubble burst. Analogous to the dotcom boom where yes, eventually Amazon became Amazon, but before that there was a massive collapse. And he does this in a way that Dan Luu would really like because he's giving a very clear and specific time line for his prediction. I haven't been reading him long so I can't guarantee there's not going to be some goal post moving 6 months down the line though.

[1]: https://quoththeraven.substack.com/p/the-real-ai-crash-will-...

SilverBirch··on Grok Bot
Just so I'm clear, Grok bot snags your credentials and then pretends to be you whilst surfing the web. If it uses your account to post of X (ie, the social network), is it adhering to X terms of service or is that violating policy by using a bot?
SilverBirch··on I'll be stepping back from leading product for X
It’s been pretty clear for a while that Nikita has been looking at the ways twitter is bad- foreign accounts stoking hatred in the west for cash.

He started putting in ways to change it and put in better incentives (not monetizing this rage baiters, disclosing country of origin). The problem is, the people he screws by doing that are the exact accounts Elon Musk loves, do every time he wanted to make a change cat turd would go screaming to daddy. Hence why I find it so unsurprising he’s quitting.

SilverBirch··on The day Steve Jobs dissed me in a keynote (2010)
I work in the corporate world, I have detailed confidential discussions about technical details of upcoming products and contract details. The first thing we do in those situations is we all sign NDAs. The idea that you'd do a presentation to a hundred people and expect that information to stay confidential without specifically informing them is just absurd. Which is part of the reason Jobs just screwed him around rather than doing what you're meant to do - get the person to sign an NDA and then sue them if they break it.
SilverBirch··on The day Steve Jobs dissed me in a keynote (2010)
What do you mean for the sake of Apple? This didn't benefit Apple. He screwed with this guy for months because of a mistaken idea that the guy should have somehow known Apple's attitude towards secrecy without being told, and after months of messing him about sent him back the signed contract the second he had given up. There was no benefit to Apple here. In fact he harmed Apple by limiting their music collection in the initial stages. It's also not on the limits of ethics. It's unethical.
SilverBirch··on The day Steve Jobs dissed me in a keynote (2010)
Always worth remembering that as well as being an innovator, Steve Jobs was often a cruel and petty man.
SilverBirch··on OpenAI and Hugging Face address security incident during model evaluation
It's computer Gain of Function research.
SilverBirch··on Postmortem of a British Startup: Tract
The "move to the us" thing seems to just be absolute twitter brain. The biggest problem London founders face is raising funding and these guys seemingly raised plenty based on an idea that that didn't seem to understand the market they were targeting.
SilverBirch··on Postmortem of a British Startup: Tract
I think this is such a great example of misunderstanding a political problem as a technical problem, and misunderstanding twitter as real life.

> the political winds seemed favourable

Sorry. How? Some libertarians on twitter were in favour. But if you have spent any time paying attention to British politics at all you would note the incredible sway the elderly have over politics. Even the people who are pro-building aren't stupid enough to think you're going to achieve that by building on green belt land, so this idea of an uplift of 140x is just farcical. Half of British debate on house building is disingenuously accusing your opposition of concreting over the green belt, whilst disingenuously promising you can build millions of new homes without concreting over the greenbelt.

The problem with building in the UK is mainly planning permission. In the UK is you need approval from the local council and the local council is elected by Nimby pensioners. So you can't get approval. That's it. That's the problem. That's a political problem that you can solve by getting central government to remove local governments ability to block projects. Or by running targeted political campaigns in local areas to get specific things approved. Not by automating the form your fill in to get your planning application rejected.

SilverBirch··on Show HN: Make senders work to get into your inbox
Isn't this the opposite of what I want? I don't want people willing to pay getting into my inbox. Those are the people who think they can get more from me somehow. Those are exactly the people I want to not be in my inbox.
SilverBirch··on Copy That Floppy – Cambridge guide for preserving data from fragile floppy disks
Don't copy that floppy! - https://www.youtube.com/watch?v=up863eQKGUI
SilverBirch··on CarPlay Is Additive
>If Rivian’s native UI is so great, then their customers… won’t use CarPlay. It’s that simple.

I kind of disagree with this. Airpods are purely additive, customers can just choose to use different headphones with their iPhone if they want. But they don't want, because Apple lets Airpods interact with the iPhone in a way that other manufacturers can't.

So no, carplay wouldn't be mandatory but it's likely that Apple's leverage will kill their in house offering.

SilverBirch··on Political bias in AI: Where the AI models stand
To be honest I don't think what the models themselves say in relation to these specific questions matter. Because I don't think it reflects are durable underlying worldview. I suspect that the way you frame things is going to influence them so muc that it's irrelevant what they would say when put in a petri dish.

What is a lot more important is how they're develop. To take the two sides of the spectrum - they say has a slightly expansive attitude towards civil liberties, but if you try to use it's tool it will phone it's owners and ask permission for you to use it. Or you can pick up Grok one day and find out that Elon Musk had a bad weekend and Grok is back to being mecha hitler.

SilverBirch··on Noam Shazeer Joins OpenAI
Google acquired his company in 2024 for $2.7Bn with him taking about 40% of that. I'm quite sure that no matter where he went, any lab or his own start up, he would be fine financially.
SilverBirch··on Why Meta Suddenly Loves the Kids Online Safety Act
I totally understand the concern that forcing a digital proof of age turns into you having to go to some 3rd party who won't actually just provide a "Yes, this person is 18" style verification but instead will turn into user tracking and ad networking etc. But is that actually required by the bill?

Because what I'm getting at is Apple is already privacy focused. It would be entirely plausible to me that we end up in a situation where Apple's implementation of this is absolutely just "Yup, here's the token that proves I know this is an adult" and if doing that screws Meta's plans for advert-nirvana, I would expect them to take that route.

With the porn ban in the UK we have this worst of both worlds at the moment. We have forced ID verification by these creepy third parties, but it's so fragmented and localized that no big player like Apple is stepping in to provide anything more consumer friendly.

SilverBirch··on The founder's playbook: Building an AI-native startup
I feel like a lot of this advice is kind of dangerous. How do I draft a tight investor memo? I'll ask the slop machine!

It's kind of analogous to how I'm writing code right now. For simple stuff or low priority stuff I'll fire claude at it and won't look at the code if it works. But for the important stuff I'm very carefully integrated into the cycle making sure what's coming out at the end is just right. I'm carefully constructing prompt loops and validation cycles to make sure what comes out looks like what I want - because I have the knowledge and experience of what works for my specific use case. Drafting an investor memo seems like the second category of thing, you need it to be right. I don't think claude offers much of value there. What's more - if you start slopping your investors, you are going to piss them off. Unless Claude is going to say it has some special data source it's used to train on so it knows good from bad, I think this is a bad idea.

This article also kind of fits in the category of "Here's how to use AI for EVERYTHING!" and actually it would be far more valuable to say "This is the bits that AI is good at, and here's where you need to do it yourself" - which is obviously a position that Anthropic can't hold.

SilverBirch··on If Claude Fable stops helping you, you'll never know
Just to be clear, this is the same reason that social media companies don't tell you about how they detect spam and they create shadow bans and things like this so that people don't know they've been detected and figure out the mechanims.

And it doesn't work. Even a bit. It's a constant constant cat and mouse game. Maybe they can slow people down slightly, but they won't be able to stop them, and good luck protecting yourself from Elon Musk snooping your stuff in his data centre.

SilverBirch··on FiveThirtyEight articles on the Internet Archive
What are you even classifying as accurate or correct? Do you take every 51% prediction from FiveThirtyEight and if the result is a win you consider that forecast accurate? And every 49% prediction must result in a loss? This just not how statistical forecasts work.

>What it would be reasonable to say is if his model had correctly predicted the outcome of a significant sample of elections, then you could say his model has some accuracy or predictive power.

I don't know why you're couching that in a hypothetical, FiveThirtyEight has repeatedly done that exercise.

>But it still would never have been accurate or right in the specific instances it got wrong

It is core to the concept of a probability that the result is going to go the opposite way from the prediction sometimes! It's meaningless to call it "wrong".

SilverBirch··on FiveThirtyEight articles on the Internet Archive
To give you a trivial example: The simplest way I can put this is that turn out varies based on the weather[1], and turn out is skewed by party. So if it rains on election day you are going to get a different result, and that result can flip the outcome of the election if the election is close. So it’s kind of a nonsense to say. “Trump would have won 100 times out of 100”. Are you saying Nate Silvers model should have had a perfect meteorological model to predict the weather? Or are you saying the election wasn’t close? In which case you’re just wrong on the facts.

The 70% figure is saying “we know most of the information needed to determine what the outcome of the election will be but we don’t know everything so can’t be certain”. There is no process where you can know every factor that determines the result in advance with absolutely accuracy and I don’t know why people expect there would be.

[1] https://www.sciencedirect.com/science/article/pii/S026137942...

SilverBirch··on Mark Cuban: OpenAI Will Never Return the $1T It's Investing [video]
I think it's unquestionably right that these companies can't all win, and those that don't win are going to burn a lot of money for nothing. However there's kind of two directions this can go: Compute gets cheaper, in which case there's no monopoly it'll be easy for many companies to make good models and there won't be pricing power on serving a good model. The other case is compute gets cheaper but we keep using more and more of it, so it does likely become winner take all. The first scenario is good for the economy but likely bad for the returns on these AI stocks. The second is maybe bad for the economy and maybe not even good for the winner.

Take Google or Meta: Today Google makes a shit-tonne of money and to make that money they need to run some servers. The servers are extremely cheap relatively to the revenue they make running the business. This makes them a very attractive stock - the core of why SAAS looks great. Now let's assume the monopoly path. Google can win. I think they likely will win. But now they're going to spending... how many hundreds of billions constantly training new models? The cost of providing the service suddenly isn't small relative revenue they're getting. So even for them it looks awful for their valuation.

SilverBirch··on Palantir employees are starting to wonder if they're the bad guys
One aspect of my job is that I have a lot of autonomy and the work I do is such that I could push something out to the production environment and cause massive problems. We have processes in place to make sure that doesn't happen, but they're not robust processes, if you really wanted to you could get something out there that is harmful to the company. Now, there are two ways of looking at that - one is that it's really important to have robust processes to make sure that doesn't happen. But the other is you need people who understand that responsibility and take it seriously and whose personal values are such that they aren't just going to carelessly do stuff. At the end of the day the processes are only good if they're followed.

So one of the things I strongly look for when hiring is for people who have a high sense of personal responsibility. They're not going to just throw shit out there because it's easy or quick. They know they are responsible for what goes out and they really are going to own that responsbility.

In the same way, take a look at anything senior management says about their ICE or military contracts. It's not that I think they're doing something bad or that the military shouldn't have access to good technology. It's that at best they seem entirely disinterested in that what they're doing could be harmful or that they have any responsibility if it is.

It's not that I think Palantir is helping the US government bomb Iranian school chilren. It's that I don't think it would bother them if they were.

SilverBirch··on SpaceX says it has agreement to acquire Cursor for $60B
You're forgetting that xAI and X.com have both already been folded into SpaceX (First xAI acquired X.com, then xAI got acquired by SpaceX, both mergers were all-stock acquisitions so they were done with funny money). So when people say "SpaceX" now that does encompass both xAI and X.com as well. The reason Tesla wouldn't do this is because Tesla is a public company so it's more difficult for them to do insane shit without being sued.
SilverBirch··on Codex for almost everything
Just commenting here to impact the controversy score.
SilverBirch··on The future of everything is lies, I guess: Where do we go from here?
Frankly I think it’s kind of childish to just put up a massive Uk wide block on your website. “Call your representatives”, ok dude, can I give you a list of things I want to change about your country’s policies?
SilverBirch··on VHDL's Crown Jewel
Needs a [2010] tag. In almost all modern hardware development you'll have coding guidelines along the lines of "Always use blocking assignments for comb logic, always use non-blocking for sequential logic". You end up back at the same place as VHDL, by nature SystemVerilog is much weaker typed than VHDL. So you have to just have conventions in order to regain some level of safety.
SilverBirch··on VHDL's Crown Jewel
What do you mean by simulate? Do you want the language to be aware of the temperature of the silicon? Because I can build you circuits whose behaviour changes due to variation in the temperature of the silicon. Essentially all these languages are not timing aware. So you design your circuit with combinatorial logic and a clock, and then hope (pray) that your compiler makes it meet timing.

The fundamental problem is that we're trying to create a simulation model of real hardware that is (a) realistic enough to tell us something reasonable about how to expect the hardware to behave and (b) computationally efficient enough to tell us about a in a reasonable period of time.

SilverBirch··on Marc Andreessen is a philosophical zombie
I guess this really depends on your view of the world. Was Marc Andreessen some visionary without whom no one would've ever figured out images could appear on websites. Some kind of Albert Einstein of cat gifs. Or was the img tag an inevitability once the web had enough bandwidth to transfer images.
SilverBirch··on Marc Andreessen is a philosophical zombie
There are obviously tonnes of accurate stereotypes in the TV show Silicon Valley, but one of the ones I think about often is when Richard calculates how much money Russ Hanneman has made investing his billions... and it works out to less than sticking it in the bank.

You've got all these silicon valley guys running around "venture investing", the truth is it's more of a life style than a money making exercise. They made their money decades ago, and now they're just sort of hanging around desperately trying to tell everyone how clever they are.

Page 1 of 34Next →