OpenAI created a team to control 'superintelligent' AI – then let it wither
techcrunch.com
techcrunch.com
I know comparing AGI development to nuclear weapon development is a trope at this point, and in general, not that that useful of comparison... But I do think having a group with a diversity of how to think about the impact of desired outcomes is a good thing to have whenever you are working in the realm of world changing outcomes, and it seems like the changes at openAI are eroding that to some extent, even if it is really really unlikely to come out of this specific company at this time.
Also, competition search engine style.. this feels is getting ahead and staying ahead.
I've decided anyone concerned about these issues knows almost nothing about computability theory so their theories are either nonsensical or just outright crazy. Very few understand the required formal concepts to have any useful ideas about how computers should be programmed to prevent "unsafe" results (which is often left just as ill-defined as most everything on AI safety and alignment research).
The problem is that OpenAI's software, especially GPT-4o, is primed for dangerous misuse, and the demo videos of GPT-4o seemed "misaligned" with any reasonable standards of AI safety. Even if Leike/Sutskever have delusions of grandeur about AGI, at least they cared about the idea of AI safety. It seems like pushing the team out meant getting rid of a lot of internal critics (and implicitly threatening anyone else who might speak up).
Rat, dog, human, bee. Doesn't matter. The moment AI is able to incrementally improve a bee intelligence will evolve into rat, blink again and its dog, another blink and its smarter than us.
That is the danger. The issue of control and danger is not linear but exponential. And without a plan to deal with issues things can get out of control exponentially faster.
> The problem is that OpenAI's software, especially GPT-4o, is primed for dangerous misuse, and the demo videos of GPT-4o seemed "misaligned" with any reasonable standards of AI safety.
I forgot who said it, maybe Sam Altman. The reason for releasing the current models was to show public what the current AI can do and how far it progressed. To get people to understand where we are and where we can get to with AI
Sort of like giving people muskets so they adapt to it before machine-guns appear.
It forces other AI research to also show their stuff and not keep everything secret, until one day out of the blue we have pocket nukes available at everycorner for low low price of $10
https://i.imgur.com/z83umbk.jpeg
Here I change a widely known riddle to the opposite answer, and I manage to make it state them both as the answer.
You can stump a person with a riddle or a logic puzzle or an optical illusion.
The fact is that people who say 'it is just fancy autocomplete' are using a thought-terminating cliche, and a 'I stumped an LLM' proves nothing.
It is simply regurgitating this phrase without even considering that it is stating the exact opposite of the answer it just gave, simply because most answers to this riddle on the internet say this at the end.
> This riddle plays on the assumption that a surgeon is typically male, but in this case, the surgeon is the boy's mother.
So from this 1 failing, you can see that it is a copy and paste machine, and it doesn't even understand that it is contradicting itself.
It's very much feels to be gearing towards figuring out the gist of a search engine you may be trying to complete and put together by reading a few links.
When I speak colloquially, I have an underlying idea rooted in a world model to be expressed. I don't spit out 1 word at a time based on the previous words I already said.
No, it "can't be debated," it is clearly false! You said "by definition," but you used an irrational and bigoted definition of "general reasoning and logic" which conflates such things with performance on a standardized test. Humans aren't innately good at stupid logic puzzles that LLMs might get a 71st percentile in. Our brains are not actually designed to solve decontextualized riddles. That's a specialized skill which can be practiced. It's depressing enough when people claim IQ tests are actually good measures of human intelligence, despite overwhelming evidence to the contrary. But now, by even worse reasoning, we have people saying a computer is smarter than "average humans." (MTurk average humans? Undergrads? Who cares!) The complete lack of skepticism and scientific thinking on display by many AI developers/evangelists is just plain depressing.
Let me add that a truly humiliating number of those """general reasoning""" LLM benchmarks are fucking multiple choice questions! Not all of them, but a lot. ML critics have been complaining since ~2017 (BERT) that LLMs pick up on spurious statistical correlations in benchmarks but fail badly in real-world examples that use slightly different language. Using a multiple choice test is simply dishonest, like a middle finger to scientific criticism.
SolidGoldMagikarp is the canary in the coal mine if you doubt it's Reddit
There are good arguments in the literature for why you might want to care about these risks [1, 2], and I think there's lots of room for reasonable disagreement about whether these are arguments are any good, but pretending the entire field of AI Safety is just delusional is just bad faith at this point. Especially when companies like OpenAI, Anthropic or GDM were explicitly created to build AGI, and have been talking about these risks since they were first founded.
[1]: An introductory paper I like is The Alignment Problem for a Deep Learning Perspective, from Ngo et al, https://openreview.net/forum?id=fh8EYKFKns.
[2]: A broader, less technical introduction to AI Safety that I like is Hendricks et al's An Overview of Catastrophic AI Risks, https://arxiv.org/abs/2306.12001
The first three risks are completely reasonable and people should be thinking about them. No, ChatGPT should not be diagnosing patients and giving them medicine. Yes, we should be vigilant to a flood of disinformation and revenge porn made possible by AI generated content.
But when people talk about “AI safety” in this context, it’s usually in reference to the fourth category, planning for a superintelligent malicious AI that evades detection, self-replicates, etc. That’s pure science fiction at that point, and it’s not a reason to slow down development of LLMs, which yes are basically glorified chatbots and will not lead to “AGI” in this threatening sense.
If I recall correctly, when steam engines started being able to go 40-50 MPH, there were people who were concerned that human beings would not be able to survive travel at such speeds because we never had experienced them. This wasn’t completely irrational, I suppose, as there are speed-induced G forces that are fatal, and they had no way of knowing the threshold back then. But once it was clear that steam locomotives weren’t in any danger of putting us over that threshold, incessant worry about locomotive-induced speeds death was kooky. “Locomotive safety” involving derailment mitigation, track crossing markings, etc. - still legitimate. But if “locomotive safety” was associated with people making claims like “we’re headed for a mass casualty event when the first locomotive hits 60 mph,” then “locomotive safety” would be marginalized.
It doesn’t help that the public faces of “AI safety” include autodidactic pseudointellectuals, clearly mentally unwell people, and philosophers too deep in their own “taken to its logical conclusion…” thought experiments.
Ok, that seems ridiculous and infeasible and nothing that has happened so far indicates that this is a real threat.
>Yes, but it would be so catastrophic to humanity, that we must take it seriously, regardless of how improbable it is!
This is the AGI debate, in my view. If we're picking different extremely improbable events to get worked up about, why stop at AI?
I'm not sure what you would do now that is different because you are anticipating such a thing.
Actually the video addresses this point, the comparison it makes is waiting until we’re already on Mars before we start thinking about spacesuits and airlocks
This is just another in a long line of technology panics. Unfortunately, there always seems to be some "concerned" experts that are both overly optimistic about the speed of technological progress and overly pessimistic about where that progress will lead adding fuel to the fire.
Eventually some other new technology will incite a new panic and this one will become another footnote in history like the fears over grey goo or genetically engineered superhumans.
My point is that the only thing that seems to convince the pro-AGI crowd is having them make specific predictions, waiting until the deadline, then asking them why they didn't happen yet.
[1]: You might disagree with this assessment, but my point is that this is not where most people are arguing from.
The parts and models are off the shelf now, it just needs the first person desperate enough to release a swarm of killcopters with cheap ML classifiers to make it a reality.
Curious how people will feel after the first autonomous swarm becomes widely known, because it will be replicated widely.
So…
The team that was supposed to control and align super-human intelligence was unable to align merely human intelligence to give them compute?
Maybe that was part of the test.
I find these theories to be extremely convoluted and implausible, and they often lack awareness of the history behind companies like OAI, Anthropic or GDM
OpenAI realised that it is not getting closer to AGI anytime in the near future and it has more significant worries closer to the present. Including existential threats. Anyone worries about dangers to humanity should focus on climate change and jump off the sci fi hype train.
Lots more discussion: https://news.ycombinator.com/item?id=40390831
https://news.ycombinator.com/item?id=40391382
And Jan's post among others:
> OpenAI is shouldering an enormous responsibility on behalf of all of humanity.
To me this sounds delusional. It assumes all kinds of things, but primarily that openai will be the leader in this space up to and beyond smarter than human agi. This self important bs is also why I wouldn't trust open ai with this responsibility
"International Scientific Report on the Safety of Advanced AI"
https://news.ycombinator.com/item?id=40400438
36 points | 53 comments
He may really be a tech version of George Santos. A grifter like that should be nowhere near in charge of tech like this.
This is to say, the objective of the Superalignment team was precisely to work on techniques that would work for models which don't yet exist. They are of course aware that they don't yet have superintelligence.
[1]: This paper by Anthropic is a good introduction to the problem; https://arxiv.org/abs/2211.03540
[2]: See, for example, Jan Leike's talk on this; https://www.youtube.com/watch?v=BtnvfVc8z8o
Who needs a safety team to control that?
Reality differs from this opinion
What makes this different? What do we actually have, stripping away all of the "it will be this someday" thinking? Pretty good chat bots, ok summarizers, things that can messily and somewhat unpredictably code at a low level, and worst of all, things that sometimes just make up what they're saying or otherwise fail in harmful ways.
Then compare that to the cost to run these mediocre miracles.
I can speak to a few restaurant owners who no longer need devs because an LLM can keep their website menu up to date. (Bonus: it’s now just a PDF.)
That’s a tangible productivity change. Granted, it’s on par with crypto’s remittance-efficiency pitch. But unlike crypto, these advantages are growing. Speaking personally, Kagi’s AI search is far superior to a list of links in most cases.
It doesn't, but that's cynicism, not realism.
Probably they even had something to say about how they legally download public data and videos.