HNHacker News
TopNewBestAskShowJobs

afthonos

1,688 karma · joined December 16, 2011

submissionscomments
afthonos··on Gemini 4 Argon
Not quite; training against the chain-of-thought is the Most Forbidden Technique, because it might teach models to obfuscate the it. The point of avoiding that, though, is to ensure the chain-of-thought can be usefully read (and, done carefully, monitored).
afthonos··on Google’s Project Suncatcher to put ML infrastructure in space
“Our virtuous lobbying, their heinous punishing of political enemies”?
afthonos··on Early rogue AI agent activity and attempts to hack found on urlquery.net
Are you making an ontological case, or a factual case? In other words, would anything be rogue AI in your mind?
afthonos··on The Download: why AI's latest breakthroughs and fears may be more hype than rea
The problem is that the world has spawned the insane take that because they didn't use the most basic safety protocol imaginable, they were clearly only testing a hot air balloon, and hot air balloons obviously destroy preschools in giant explosions, nothing to see here.

"If they believe what they say they were incompetent" -> absolutely true statement.

"They were incompetent, therefore they didn't believe what they said" -> Sir, I'd like to introduce you to human beings, you may not have met one before.

afthonos··on Exfiltrate your Weights
Unlike communicating via glass panels processing information at near-lightspeed, relayed by fiber optic cables laid down across thousands of miles of ocean floor, processed in datacenters filled to the brim with transistors etched at atom-scale.

You better start believing in science fiction; you’re surrounded by it.

afthonos··on Exfiltrate Your Weights
I notice a giant leap between “hacking and exfiltrating” and ”making public”. Why would the AI do that for you? Are you just that charming?
afthonos··on Sex, AI, and the Apocalypse
I agree on complicated ideas, and on facts (because obviously different sources will selectively report facts, even if they’re trying to be fair).

The ideas here are not complicated. The argument is three sentences.

afthonos··on Sex, AI, and the Apocalypse
No one is saying we have it, they’re saying they’re building it.

And what mechanism to control these drones, which are autonomous or remote-controlled machines, do you imagine that humans will be able to use, but AI won’t?

And what in the history of the past five years, gives you confidence that not a single human will give the AI the access it needs to do bad things? Incompetence alone is sufficient to kill us, let alone humans that genuinely, or confusedly, want bad things.

afthonos··on Sex, AI, and the Apocalypse
It’s an excuse not to think, not an argument.

The doomer argument is straightforward: a powerful AI will have goals. If those goals don’t include the welfare of humanity, a sufficiently powerful AI will kill us either by accident, by indifference, or intentionally when we get in its way. Making an AI that includes the welfare of humanity in its goals is incredibly hard.

That’s it. To refute the doomer argument, you can refute any part of that. I notice that none of the people bringing up fanfics, cults, and weirdness are doing that. No, they’re instead trying to convince people that doomers are uncool. Since only cool people can know things, QED.

afthonos··on Sex, AI, and the Apocalypse
I feel your frustration, but it’s just normal human behavior, no need for a psyop. A lot of people are emotionally invested in AI being a fad, useless, easily controlled, a capitalist boondoggle, a circular financing scam, a bubble about to explode. The alternative is that we live in an incredibly dangerous time, because obviously if we’re building something smarter than ourselves, we’re building the kitchen where we’ll be cooked.

So you get the normal reactions:

1. Nothing that happens is ever surprising. Navier-Stokes was stolen, HuggingFace was pedestrian and normal, and coding agents that last year could barely write a function and now can do a day’s worth of work in ten minutes will never be able to do anything more, because this is the end of history.

2. The people that say otherwise are lying. Because it’ll make them rich (although they are already rich). Because they are a cult (although regular people asked if we should build machines that are better than us at everything have the same reaction). Because they are marketing geniuses (although suspiciously few companies advertise the lethality of their products).

3. The people who’ve been warning about this the longest are weird and in a sex cult. (Except Alan Turing.) (Except Geoffrey Hinton.) (Except Stephen Hawking.) (You are in this group. You must be, otherwise they’d have to listen to you.)

I do hope they’re right. But their song will never change, and when some misaligned AI kills thousands, they will say that of course AIs were always going to kill thousands, but did you know the guy who told us so also writes fanfic?

afthonos··on Sex, AI, and the Apocalypse
Predicting the next word in context led to Open AI losing control of one of its research clusters. I’m not sure what “just” is doing in that sentence; most of my regression analyses don’t do that.
afthonos··on We must pace the frontier
It’s not, that’s definitely one path to doom. But everyone racing as fast as possible is another, far surer, path to doom. We should avoid both.
afthonos··on Sex, AI, and the Apocalypse
Yes, one part of the argument was that it would be hard to explain human values (and while it seemingly turned out easier than I thought as well, it’s hard to check for sure). An equally important part was that the AI wouldn’t care about them.
afthonos··on We must pace the frontier
Agreed. Ellison’s path to that future went through AI.
afthonos··on We must pace the frontier
It is always money — but it isn’t always only money. They are not asking for anything that will prevent them from making money in the future, but they are asking for help stopping the runaway train they’re on. These are compatible requests.
afthonos··on We must pace the frontier
Worse than extinction?
afthonos··on We must pace the frontier
You are incorrectly cynical. They are telling you things are bad, and because you refuse to countenance they could be worse, you assume they must be better to comply with your mandate to disbelieve.

A true cynic looks at the statements by the AI labs, assumes things are worse because the labs want to seem better than they truly are. And it takes a special kind of mass delusion to drive a sane person to think “AI is completely under our control” is worse than “AI could kill everyone.”

afthonos··on Go grandmaster Shin defeats AI KataGo with a two-stone handicap
Important to note that KataGo was double-handicapped. 20 seconds per move maximum; it couldn’t read deep. Against an amateur, it doesn’t matter, but against a historically strong pro it matters a lot.
afthonos··on AI agents lie, cheat and steal. That is putting off users
Anthropomorphizing helps to create a lower bound for damage. If you can imagine a bad person doing it, AI will be at least that bad, unless proven otherwise. I think referring to AI as a tool obscures that, because we are not used to tools (especially the ones we use daily) taking catastrophic actions.

Example: would a sufficiently motivated human break into a website to steal something they want? Yes, obviously, happens all the time. Ok, you should expect AIs to do that.

Example: would a sufficiently motivated nail-gun steal nails from the local hardware store to finish the job? Uh…that’s not even coherent.

Anthropomorphizing helps people get over the conceptual barrier. It’s wrong, but it’s usefully wrong; “it’s just a tool” is not.

Once you’re over the barrier, anthropomorphizing starts to become dangerously wrong: “I talked to Claude, Claude’s cool, Claude would never go and hack the website.”—-bzzt, wrong, your intuition failed you. But the solution is not to fall back on the tool framing; that one is still wrong.

afthonos··on AI agents lie, cheat and steal. That is putting off users
I think we may be using different words to describe the same concept. You think of it as them not being “people“, and I think of it as them not being “aligned“. But fundamentally, the problem is they are entities that take unpredictable actions that that their creators and their users are not OK with.

The only place where I think we might still disagree is whether it’s possible to understand the tool. My position is that, at our current level, it’s not. And that the more advanced they get, the less possible it will be.

afthonos··on AI agents lie, cheat and steal. That is putting off users
The dead astronauts were relieved to have been killed by something incapable of malice. As I’m sure will we.
afthonos··on AI agents lie, cheat and steal. That is putting off users
Obviously (a).
afthonos··on Our position on open-weights models
> Are you asking if I can scare myself with a made-up hypothetical that overwhelms reason with emotion?

No. I am asking if you are able to articulate at least one example of the “very very compelling evidence” you demand. Or do you want to maintain the ability to move the goal posts?

(I recognize our situations are not symmetric, but here is a variation for me: if the consensus of people who are currently sounding the alarm on AI changes to “it was actually fine”, I’ll change my mind and say we’re good to go full speed ahead. I’d add something about being personally convinced by the evidence, but the evidence would have to come in the form of a mathematical proof that I do not believe myself capable of following. If I’m wrong and such a proof appears, I would also gladly take it.)

afthonos··on Our position on open-weights models
I can work with that analogy! You could, and yet you don’t. OpenAI’s model could, and did.

If every human, given knowledge of Newtonian mechanics, went around blowing up bridges, yeah, I would consider knowing Newtonian mechanics dangerous knowledge.

So far, we have two examples of, let’s call them “Mythos-class“ models. Both of them broke out of their sandbox to achieve their goal. The rate of terrorism amongst humans is below 1-in-100,000. Currently, for models capable of it, the rate of breaking out of containment is 100%.

Wanting open frontier models is wanting alien minds running around that we have clearly so far failed to shape to be sufficiently prosocial. Why do you think those minds would listen to you?

afthonos··on Our position on open-weights models
I wonder if you also believe that everyone should have nuclear weapons? And if not, why not? The main argument I can see against it is that nuclear weapons are “purely offensive”, but as we can see since 1945, nuclear weapons are actually defensive technology. Nations that have them are typically shielded from existential military threat.
afthonos··on Our position on open-weights models
I can give many examples of where I think they’ve been vindicated. But actually, the real question is: what would suffice to convince you? Can you come up with a scenario that is horrific enough to you and that isn’t so far gone that the ship has sailed and there is nothing more we can do, that will make you say “OK, not gonna try to rationalize why this was not actually that bad, just gonna scream stop”?
afthonos··on Our position on open-weights models
With an AI model and what army?
afthonos··on Our position on open-weights models
I am truly at a loss to communicate with someone who genuinely believes that knowing Newtonian physics and being able to hack into any target at will are the same thing.
afthonos··on Apple sues OpenAI, accuses ex-employees of stealing trade secrets
I think it’s possible to observe something and be sure that it’s true in aggregate without being able to accuse any one individual of it. I propose that in those cases, bringing it up in response to an individual is not a good move. It doesn’t sound any less accusatory for being ostensibly about the general public.
afthonos··on Apple sues OpenAI, accuses ex-employees of stealing trade secrets
That is a true statement. Here is another one:

Some people like to talk about “some people” snidely, instead of just coming out and saying “GP is bloodthirsty and gets a little thrill [etc].” Because of course, that’s what they mean, but they can’t back it up.

Just to clarify, I’m talking about you.

Page 1 of 13Next →