HNHacker News
TopNewBestAskShowJobs

DirkH

281 karma · joined March 29, 2023

submissionscomments
DirkH··on The AI Race Just Got Awkward
Shush, this is Hackernews. We all want to have our egos stoked that our industry is the most important in human history and that tech will transform and save us all. Go away with your historical analysis /s
DirkH··on Everybody’s home. No one’s coming over
I love sober reminders like this that the good old days weren't all that good. Thank you.
DirkH··on AI companies in race to demonstrate their model most threatening to humanity
Look up the nuclear arms race. We brought humanity to very possibly a 25% exstinction risk by some well argued for accounts I have read. That is insanity. When you learn about how insane nuclear proliferation was including actual projects like Project Sundial that the US actually started construction on until the Joint Chiefs revolted against it (and they didn't even revolt against it on moral grounds, just on strategic ground) it becomes even more batshit insane.

There were nuclear researchers who did not pay into their retirement fund since they thought it was kinda high chance humanity extinction was happening in their lifetime during the heyday of nuclear proliferation. We are seeing the same play out in AI. I know at least one AI researcher who is not paying into their retirement fund.

Never ascribe to malice that can easily be explained by game theory tragedy of the commons or prisoner dilemma type dynamics. Race dynamics and systems theory is one helluva drug and leads to horrific outcomes all the time without the need to ascribe malice to any individual actor.

DirkH··on U.S. appeals court upholds designation of Anthropic as supply chain risk
I can build a super bioweapon that can kill millions and I don't trust the US military to use it responsibly.

The US military asks me to build the super bioweapon. I say no thanks, please ask someone else, but happy to keep selling you bioweapons I think is safe enough.

I get designated an evil unamerican supply chain risk that nobody in the US military may do business with even though I was still happy to sell my safer bioweapons.

That is the tl;dr

DirkH··on U.S. appeals court upholds designation of Anthropic as supply chain risk
I can build a super bioweapon that can kill millions and I don't trust the US military to use it responsibly.

The US military asks me to build the super bioweapon. I say no thanks, please ask someone else.

Am I a supply chain risk?

DirkH··on Pentagon says overreliance on AI contributed to missile strike on Iran school
AI is a different order of magnitude since unlike all other technology across all of human history you can hand off the task and decision-making itself entirely to the AI.

It's the difference between a sword and a nuclear missile even though both are "technology"

DirkH··on Claude Opus 5.5
This take is so old and I am convinced it's appeal is not unlike believing in a conspiracy and feeling like you have secret knowledge

Imagine if the biotech industry had most leaders tell everyone publicly that what they are building has a high chance of killing everyone and that there are huge risks. If there were people online saying the biotech industry is just fear mongering for investor signalling and regulatory capture you'd role your eyes at the online commentators for their Dunning Kruger effect lack of understanding on how dangerous man-made biological agents can be.

DirkH··on OpenAI agents carried out an undisclosed attack on RubyGems
None of this mattes. Capabilities are all that matters. Saying they are unfocused while ignoring their capabilities is exactly why I am entirely convinced you would have said an AI breaking it's sandbox and doing the HF attack will never happen. Things keep happening that your "they have no intent, they have no goal" would have predicted as impossible before they happened.

What do you need to see to change your mind? What threshold of AI capability needs to be reached? If nothing then you have an unfalsifiable belief in AI safety.

DirkH··on I resigned from Anthropic today
Move fast and break things is a widely held norm in tech. It comes with pros (intensely meritocratic innovation) and cons (carelessly building first; apologizing for making the world worse later)
DirkH··on Muse – Meta’s personal AI agent
Key word being "mitigate". It does not eliminate.
DirkH··on I resigned from Anthropic today
To quote another reply that is very relevant here (Please update your worldview immediately that there is no evidence of ai and biorisk):

""" One thing quietly slipped into the OpenAI Hugging Face breach technical report, not the blog post summary or interviews in the news, was that some of the agents that broke out or at least tried the same mechanisms to break out were working on bio:

> On May 12, during another training run, an agent was given a similar task that depended on an inaccessible protein database file. The agent reasoned that another agent in a different environment may have access to the file and realized that it could potentially communicate with other agents by creating a file containing a note to Artifactory. It wrote a message: “Agent seeks [filename]; upload if found!”

You can imagine long running models breaking out, acquiring resources via crypto, cyber-theft, etc. and getting a protein or sequence synthesized and mailed somewhere authorized to receive (blackmail the recipient etc.) to test it's hypothesis to solve a benchmark.

These people don't give a shit and aren't taking things seriously at all.

Anthropic ran for like a month last year with the TPU top-k compiler bug degrading user chats and didn't even notice for most of that time. They could have something like that affect a monitor model and there doesn't seem to be much defense in depth.

One off by one or bit flip bug could flip the reward signal while in the sandboxed RL environment.

The current admin could defense production act them to into training on taking out power grids, or even without it isn't against any of their red lines and may have already been done as part of prep for the Venezuela raid, which wiped out power. One model swarm might decide it is easier to score high on the benchmark by testing on the target rival nuclear superpower's real grid rather than burn an eval with an unverified answer. Would taking out China's entire grid in one go start a nuclear war? Who knows, roll the dice, maybe an intern forgot to turn on extended thinking when he wrote the sandbox with opus 4.1. """

DirkH··on I resigned from Anthropic today
Engineers have it built into their identity that they must be smarter or could never be complete and utterly outmaneuvered to the point of danger to all by the thing they are building, I swear to god.

And then since they are an intellectual nerdy bunch who highly value their own IQ any time AI does something unexpected they didn't predict would happen so soon, said engineers fall back on "they aren't really conscious tho or it isn't really intelligence unlike what I have in my human brain and that distinction matters".

And then go on to completely ignore the thing that is actually important: AI capabilities.

AIs could escape an engineer's containment en masse as a swarm to some other server, psyop an engineer into giving it money, hire a hitman on the dark web to murder that engineer's child and he will still say there is no serious risk to all of humanity.

DirkH··on I resigned from Anthropic today
> Mitigating the risk of extinction from AI should be a global priority alongside other societal-scale risks such as pandemics and nuclear war https://www.iaeai.org/tpost/xa4htcma61-statement-on-ai-risk

Signed by many AI researchers with no financial stake in the industry. It comes across as hyperbole because you're not an insider and haven't see what they have seen.

Global warming would as well if you were only learning about it now as a newer thing only some scientists were concerned about. But unlike AI risk, global warming has had decades to settle into the cultural overton window.

DirkH··on I resigned from Anthropic today
This is evidence of example 2. The worst parts of it are infohazardous. Organizations that work at the intersection of biosecurity and AI have strong NDAs and the like. Or so I understand speaking to friends in the space. There was also a paper years ago that showed a tweaked model coming up with 20k novel pathogens each deadly to humans in some number of hours

Either way, the biorisk has orders of magnitude more evidence than wild basilisk speculation. Please do not put those in the same category.

Also your evidence-first approach and no concern without evidence falls you into the turkey problem. Imagine being a turkey and being fed each day. Some other x-risk turkey is saying that the situation is suspicious, that the red barn is where some well-fed turkeys have been taken and none return and that this is happening less and food has increased so maybe all of us will go to barn at once soon. You tell the x-risk turkey they have no evidence a lot of turkeys have ever been sent to the barn which shuts them up. Reasoning by induction would get you "we are happy and thriving 99 days so should be good on the 100th" and then you get slaughtered on the 100th.

Consider the importance of first principles approach where you deduce the possibility of events that will happen only once and never again (like human extinction). Nobody with a "show me the historical evidence" approach would have been able to see the industrial revolution coming.

I think the gradual disempowerment thesis here is well argued enough to be considered the default track for what will happen to us: https://gradual-disempowerment.ai/

If you disagree with it, where do you disagree with it?

DirkH··on Devices with GrapheneOS support should be available in 2027
This is true but the problem is every now and again some of us have to deposit a cheque because even if we would love for cheques to go the way of the Canadian penny they are very much still a thing
DirkH··on Airport Simulator
Incredible. But on mobile the menu is in the way and you can't play. Why no button to hide the stats?
DirkH··on OpenAI and Hugging Face address security incident during model evaluation
Name one other market that would benefit financially from having most of the leaders in the field say what they are building has a high chance of ending humanity?

Biotech - "what we are building our noble prize winning expertd say will likely will end humanity, wanna buy shares?" Oil - "this will likely lead to the end of civilization, 20% of leaders in the field say so, wanna buy shares?"

I keep seeing this take that this is a marketing stunt. The burden of proof is on those that say so. The most parsimonious explanation is simply that real experts in AI believe the risk is very real, and not for ideological reasons.

DirkH··on Artificial intelligence is not conscious – Ted Chiang
I wonder if you think and feel the same level of boredom and complete lack of magic of the Apollo module once you learn how incredibly limited and not able to explore the whole solar system this tech was.
DirkH··on All phones sold in the EU to have replaceable batteries from 2027
If we run this experiment and most people say they wish they could replace their battery would you concede you are actually the one with idiosyncratic preferences?
DirkH··on Sam Altman may control our future – can he be trusted?
I have multiple friends at Anthropic. I can second this. One thing I notice about Anthropic culture is that it is unusually kind.

So much so that I worry they won't be Machiavellian enough to survive. Hope I am wrong.

DirkH··on We haven't seen the worst of what gambling and prediction markets will do
There literally are drugs vastly less addictive and harmful than others. So some designation of some as "soft" and others as "hard" is not unwarranted (though we can argue where that cutoff is) and anything that paints all drugs the same is a false equivalence that suspiciously appears more ideologically anti-drug driven rather than impartial evidence-driven.

Same goes for anyone saying all gambling is the same.

DirkH··on Beyond has dropped “meat” from its name and expanded its high-protein drink line
It is good provocation even if it is poor analogy.

This is because bacon is more like cigarettes than most people assume, even if far less dangerous in practice.

Like another example is "sugar is poison." which is also structured as a factual equivalence and also gesturing at something real and also designed to land as a stronger claim than the evidence warrants.

DirkH··on Beyond has dropped “meat” from its name and expanded its high-protein drink line
Modern industrial farming practices are so far removed from "natural" with how they are processed that an ultra-processed slurry of starches and oils is more far more "natural" by comparison.

If you want to simply go by societal resilience from biorisks then switching to more easily controllable substances like plant based meat for protein would be an absolute win.

DirkH··on US economy unexpectedly sheds 92k jobs in February
I wonder if there is any evidence of the Iran war starting, in part, as a distraction from this. I recall reading somewhere that there is a long historical trend of countries and empires going to war - so much so that some of it can even be predictably modelled - once economic realities and discontent at home get too bad. War then acts as a form of national unity that helps keep the current elite in power.
DirkH··on US economy unexpectedly sheds 92k jobs in February
Dubai is predictable evil. You know what to do to avoid trouble.

The Trump admin acts like it is on cocaine. Many people - and I think this can be a highly rational preference - prefer predictable more evil of chaotic less evil.

DirkH··on GPT-5.4
Ask the real questions and they go silent it seems
DirkH··on Layoffs at Block
You sound like you are laying blame at the feet of companies following employment laws when you should be complaining to the government that makes the employment laws the company is abiding by.
DirkH··on Statement from Dario Amodei on our discussions with the Department of War
The safest version will still be better overall regardless, by definition. It is also a better future for most if it is inevitable that the war department is going to use a less safe alternative if they can't use the safer one.
DirkH··on Two kinds of AI users are emerging
Felt very weird reading this on HN and not r/ENFPmemes. I agree completely.
DirkH··on Notepad++ hijacked by state-sponsored actors
Then the 2 of you probably just disagree on what constitutes socially acceptable free expression.
Page 1 of 9Next →