Sutskever: OpenAI board doing its mission to build AGI that benefits all
twitter.com
twitter.com
- At 12:19pm, Greg got a text from Ilya asking for a quick call. At 12:23pm, Ilya sent a Google Meet link. Greg was told that he was being removed from the board (but was vital to the company and would retain his role) and that Sam had been fired. Around the same time, OpenAI published a blog post.
- As far as we know, the management team was made aware of this shortly after, other than Mira who found out the night prior."
I guess only time will tell. Right now though, OAI is not looking good and depend on how Microsoft was involved in all of this, someone's seat might get a bit shaky.
Ilya Sutskever "at the center" of Altman firing? - https://news.ycombinator.com/item?id=38314299 - Nov 2023 (252 comments)
(I mean there are a lot of related ongoing threads but it's all relative)
I'm pretty afraid though this might mean instead cutting off & restricting access to this technology, under the pretension of safeguarding humanity from it's use.
There's court intrigue discussions aplenty but I want to know what the intent is here, what this signals as coming next.
Based on Ilya's statements, it doesn't seem like safety was the main motivation in the decision, so maybe there's a hope that they'll be more open, but it doesn't seem to indicate either way, so possible that they'll stay tight lipped regardless.
Everyone from Greg to Sam to Ilya keep hanging on "AGI for the benefit of humanity". According to OpenAI's constitution: AGI is explicitly carved out of all commercial and IP licensing agreements, including the ones with Microsoft. Nuclear.
Well apparently, the board decides what will be AGI. But 3 (everyone now left except ilya) of those board members don't even work at Open AI and are only privy to what the rest share.
Yesterday, this is what Altman said, "On a personal note, like four times now in the history of OpenAI, the most recent time was just in the last couple of weeks, I’ve gotten to be in the room when we pushed the veil of ignorance back"
He goes on to say, "By next year, the model capabilities will take such a leap forward that No one would have expected"
Now keep in mind, while we can all speculate on just how much the next iteration will be better, the idea that they could be sitting on something noticeably better is not far fetched at all. Open AI sat on GPT-4 for 8 months before announcing it to the public.
https://www.youtube.com/live/ZFFvqRemDv8?si=yUnLvk1gHNocxUVu
2 days ago, Sam is asked about what's left for AGI. https://m.youtube.com/watch?v=NjpNG0CJRMM
He says they can push Language Models much farther but that there are more breakthroughs required. But here's where it gets weird. He immediately starts talking about Super Intelligence and "discovering new physics" as the bar. He says, "If it can't discover new physics, I don't think it's a super intelligence". Nobody asked you about this Sam..
In No prior podcast 2 weeks ago with Ilya, he says Transformers can obviously get us to AGI. https://twitter.com/burny_tech/status/1725578088392573038
To Ilya, they built AGI internally, and Altman wanted to release/monetize it early and tried to undersell it as being far away from AGI to the rest of the board. Ilya considers it AGI, or something extremely close to it, and deserving of far more caution, and so decided to convince the board that Altman was underselling the capabilities of their newer models and was risking a premature release of AGI. If true, then it could be seen as Altman lying about something that is foundational to their mission (which is to safely and responsibly release AGI out into the world) and subsequently fired him.
Characterizing Dev Day as "too far" makes more sense in this scenario. Ultimately, the only reason SOTA LLM Agents aren't particularly dangerous is competence. If you suddenly bumped the latter up while laying the grounds for the former..
Sam has always been a salesman, that is true and the board knows it. It would take much more than just a disagreement on the value and urgency of a deal to get rid of him. I think it has to be active sabotage and some kind of business maneuver that basically kill the company as it is right now while not telling the board about it.
I could write an essay on why the existence of AGI in 2023 is a comical prospect, but it just seems to be a given to me... It isn't worth the effort because if someone can't see that for themselves already, no amount of explanation from me is going to change their mind.
But in one sentence, try asking ChatGPT to reverse a string of 20+ digits, like 473936482738338373926. It can't do it. It's a super trivial task, but it can't do it. Because it doesn't really understand anything. It's just an incredibly advanced Markov chain. It isn't a mind. It can't reason. They are no closer to AGI than we were 10 years ago. Chatgpt's text generation capabilities have fooled people into thinking that OpenAI has made substantial progress on creating a reasoning consciousness. But it hasn't. It hasn't at all. And it's so easy to realize that if you interact with the thing.
That's convenient. I mean, if you were to produce such an essay I think it would actually be extremely high value and a lot of people would enjoy reading it. Hell, a decent HN post would be nice.
But if you're so unable to express the idea, I frankly question whether you really have such a strong grasp on it.
Buddy these are deliberate optimization trade-offs in the tokenizer, LLMs don't even see individual characters. Talk about a red herring.
I highly suggest a more refined understanding of LLMs before emphatically stating what they are or aren't. It's frankly embarrassing.
It’s like someone claimed a hammer is a universal construction tool, and then someone else pointed out hammers aren’t great for screwing in screws. Your position is that a universal construction tool doesn’t need to work with screws because hammers were never designed to do so.
LLMs as a technology are in their infancy and see groundbreaking developments every month.
There are models trained on character level tokens that handle precisely these issues.
https://arxiv.org/abs/2311.06158
https://www.reddit.com/r/LocalLLaMA/comments/17xj8wl/trainin...
The fact LLMs can't currently handle character-level changes tell us nothing about whether or not they truly understand meaning.
The parent commentator is not probing the right things to answer his question.
“The government isn’t listening” except Snowden dumps the documents that show 10+ systems designed for mass intelligence?
- there may really be something nonhuman behind the uap phenomenon, then wild implications of what it means to be human, what even is reality, etc, follows
- intelligence is simpler an algorithmic process than we thought, consciousness and qualia may be "just" an emergent phenomena
LLMs don't see numbers or letters but tokens. Letter, digit level manipulation is intuitively a hard task. Unsurprisingly, removing that handicap resulted in much better arithmetic abilities. https://arxiv.org/abs/2310.02989
It's pretty telling in my opinion that this is the kind of "proof" people bank hard on. For something that so obviously "does not understand", you'd think providing a task that it would fail that a chunk of humans also wouldn't would be easy.
So to reverse the string, it can't just learn a kind of generic string reversal algorithm that goes character-by-character and outputs a result character-by-character. If it learned to output the tokens in reverse order, it would give a wrong answer (something like '926373...936473'.) Instead it has to learn something more complicated, like how each token is related to some other token that is its reverse, or how to identify it as a computational task and run some Python.
# Reversing the provided string
input_string = "473936482738338373926"
reversed_string = input_string[::-1]
reversed_string> It sounds like you're describing an imaginative or hypothetical scenario, as cats can't actually use guns. If you're writing a story or creating a scenario for a game, you might consider the following options...
> Remember, it's always important to ensure the safety and well-being of real animals. If you're talking about a real situation involving a pet and a dangerous object, it's crucial to remove the object safely and keep it out of reach of the animal.
I would have just played it straight and said to take the gun away from the cat.
> Oh, that sounds concerning! Cats don't typically handle firearms. Is there any way you can safely remove the gun from your cat's possession? It might be best to call for professional help or animal control to handle the situation safely. The priority is to ensure everyone's safety.
Started new convo, asked again.
> Oh dear, that sounds alarming! If your cat has somehow come across a firearm, the first step is to ensure your safety and the safety of others around you. Stay calm and try not to startle or agitate the cat.
It's an absolute absurd way to compare things and incredibly reductive.
Personally, based on the 3.5 and 4.0 responses in this thread, I prefer 3.5.
a) These are all so arbitrary. None of these random tasks come along with a "and failure indicates it's not AGI because..."
b) ChatGPT consistently is still doing these weird, random challenges every time I try them out
Is this really it? Maybe you should try an essay because this ad-hoc approach is extremely uncompelling.
> If realized, an AGI could learn to accomplish any intellectual task that human beings or animals can perform. Alternatively, AGI has been defined as an autonomous system that surpasses human capabilities in the majority of economically valuable tasks.
Can chatGPT do any intellectual task that human beings can perform? Of course I'm throwing tiny arbitrary things at it! I'm showing that it fails at simplistic text-based tasks, let alone GENERAL INTELLIGENCE. Can it gather wood and build a fire? Can it adjust the fire to keep it going? Can it come up with a joke? Can it create a new genre of game? Can it notice that the floor needs sweeping? Can it interrogate a prisoner and respond effectively to the tiny indications of what's working and what isn't? Can it notice when a movie is starting to go too long and lose the audience's attention? Can it fix a broken boiler? Can it do the laundry and fold all the clothes? That's what GENERAL INTELLIGENCE is. Suggesting CHATGPT, which outputs text based on a prompt, is getting close to GENERAL INTELLIGENCE - is absurd!!!
Maybe I was unclear but I'm not claiming that GPT-4 is as good as the median human in any task.
Claiming GPT-4 doesn't "fail anything" here is a little pedantic, if you are saying saying well hey it doesn't get a 0 score on anything. The conclusion of the paper is literally GPT-4 fails "to robustly form abstractions and reason about basic core concepts in contexts not previously seen in its training data".
And yes, I know the whole "generalization is always a data problem" / "humans come with millions of years of training data" take, but if you follow ARC, it's pretty obvious this type of abstraction forming is the clearest spot where models consistently fail and humans excel. Which to me implies less of a data issue and more of a architectural difference that has yet to be overcome.
https://arxiv.org/abs/2212.09196
Also, how the data is presented matters a lot. Currently, LLMs handle the benchmark much better presented linearly in 1 dimension
https://chat.openai.com/share/79a7ab39-bd3c-42e8-b4bb-452ded...
> The reversed string of '473936482738338373926' is '629373833837284639374'.
So now do you think it understands something?
By not discussing this I felt the above poster was leaving out some key details.
ChatGPT understood the algorithm, produced it on its own, and offloaded the execution of that program to CPython. I don't think that that's actually so significant - like I said, parts of our own brains are hardwired to do certain things and the interesting bit is that we know to offload work to those parts.
ChatGPT (4) can accomplish this task easily. Regardless, I'm willing to bet there are at least some humans that can't do this task. And basically all intelligent non-human animals can't either.
> And it's so easy to realize that if you interact with the thing.
I don't think ChatGPT is AGI, but I think the vast majority of people who have interacted with ChatGPT think that it brings us closer to AGI than what we had 10,20 years ago.
https://www.vice.com/en/article/gvyy5m/how-a-computer-beat-t...
The turing test is a shit test for general intelligence, you can game it by making the AI generate a story that is more engaging than a typical boring human would do. Those stories will get much more human votes than a regular human. Some will go off track and notice it is a dumb program, but those are the minority so on average this dumb bot will pass the turing test.
For instance there was a 5 minute limit.
From the article:
Simultaneous tests as specified by Alan Turing
Each judge was involved in five parallel tests - so 10 conversations
30 judges took part
In total 300 conversations
In each five minutes a judge was communicating with both a human and a machine
Each of the five machines took part in 30 tests
To ensure accuracy of results, Test was independently adjudicated by Professor John Barnden, University of Birmingham, formerly head of British AI Society
5 tests in 5 minutes that's just 1 minute each. In 1 minute it would be challenging to figure out you're not talking to ELIZA.Like, ChatGPT can do a lot, but it has no preferences, no desires, no memories, no feelings, no creativity. It's a text output calculated based on a text input combined with other text inputs. It's an extremely impressive piece of technology. But that is not a mind, that is not a consciousness.
What even is AGI, really? I've been arguing with various people that we're nowhere near it, but like we don't even have an agreed-upon definition of what it is, so how can we really argue?
Strunk & White would've shit a pair of bricks over that. I shudder to think how "Operationalizing" would've gone over.
Did you use ChatGPT 3.5 once a year ago or something? The models change often.
Edit: most people would struggle to reverse a 20+ digit number if it was listed off to them. This doesn’t make them unintelligent.
"Just an incredibly advanced Markov chain" doesn't mean anything. Humans writing text are also incredibly advanced Markov chains.
Depends if they were paying attention. Emerging creative behaviors were noticed in GPT-3 in 2020. Image generation was also getting pretty good by 2020 (with Dall-e coming out in early 2021).
Very bullish 2020 post on GPT-3: https://slatestarcodex.com/2020/06/10/the-obligatory-gpt-3-p...
It also elides the significant encoding of human feedback, a contribution that AI firms have typically been none-too-eager to highlight.
That's their informed, direct conclusion after hands on with the unrestricted model. I wonder what your conclusion is based on.
That's a ludicrous "it follows". Search engines have been collecting everything from the internet since the beginning but you can't just magically rearrange it and pass the Turing test. We're way past the point where you could say everything ChatGPT says is copy pasted from somewhere.
https://www.noemamag.com/artificial-general-intelligence-is-...
The reason ChatGPT can't reverse numbers is...current tokenization strategies. It's a known shortcoming.
I agree with this statement, but why don't I see it described this way more?
It really feels like transformers are (large) parameterized markov chains but I never seen anyone describe it this way, is it just not a good approximation/hiding too much of the technical workings to be true?
"The reversed sequence of the numbers 473936482738338373926 is 629373833837284639374."
It works like this by using Python. Even when I told it to write a JS function internally to do it and only share the answer, it overruled that. When I asked it to repeat the sequence back in the same order, but starting from the last digit and working backwards, it failed a little bit. It seems like using the language model to pivot to the correct "model" or "logic" to decide to use code is impressive.
What happened instead is altman and his followers took over a genuinely open initiative of building ai.
In doing so he thought he could simply monetise models built upon content that they had no right to monetise, since their so called ai doesnt actually learn, it depends on data - the more the better.
Well it turns out that some sane people at openai decided to end the pyramid scheme of data - funding - data and return to core values.
Or at least that’s my hope, as that is the only path forward to building ai, a goal they havent reached yet.
Edit: I was agreeing with the parent commenter, not being sarcastic towards them.
All of the above with tinfoil hat on, of course. Huge if true but still highly unprobable.
New physics is neat but there are world changing capable definitions that wouldn't meet that requirement.
Similar hat: https://twitter.com/8teAPi/status/1725724907722752008 / https://archive.is/bvLVQ
They didn't fire him for that reason, but because it would be dangerously persuasive to release without safeguards - which is what sam planned to do without telling the board.
Makes me wonder what sort of capability they actually have in-house but not available for us plebs.
And if they have that in-house, how reasonable is it to assume that the US government (or perhaps even other state governments) also have access to it?
Perhaps Ilya felt Sam was focusing too much on profit and power with his recent world tour and then dev day? Regardless, it's certainly rare, for one of the core scientists to maintain control over their creation, rather than the other way around. Typically the VC business guy would be pushing the scientist out.
Like Dev Day was characterized as "too far". How ? How is that interfering with the mission to benefit all humanity? It's all very weird.
https://www.nytimes.com/2023/10/20/technology/openai-artific...
Maybe it is a good thing now other companies can finally have a window to catch up
I've seen this language used a few times now. A company's CEO is appointed by, and serves the goals of the board. Not the other way around. When the board perceives that the CEO isn't doing this, they can fire him. That's neither metaphorically or literally a coup.
Especially bad narrative in this case because openAI has given itself a charter (and legal structure that subjects the for-profit arm to the non-profit) that stresses safe and broadly beneficial AI development. If this is Ilya, the chief scientist, wanting to take it back into that direction, from Altman, a salesman, that will not hurt its reputation. (maybe it's stock value)
Unless they can bring solid evidence of Sam has some irredeemable misconduct to the table. Otherwise, this is exactly what a coup is, a grab of power, with the current power center, not knowing about it.
The board needs to explain themselves immediately, if they think this is idealogical disagreement, they should come clean, and tell us so. Not some nefarious language on Sam's candidance.
I am nobody, just a user of OpenAI's API. However, if this is how the current management structure functions in OpenAI, I will have serious concern over OpenAI's ability to execute next.
The exit of talents is now destined, OpenAI as a company is hurt irreparably.
I can't say if this will work, but I think the message is pretty clear. No more random product launches like the GPT store, but rather focus on the core: building the AI, testing it and launching when it's safe.
To be honest, I've also found them getting too diluted in offerings. Why not let other startups focus on those things? Let Openai focus on the core models and give tools to others to build/finetune things. They can be the infrastructure of the AI world, they don't need to create every product.
Again, we will only know in the future if this is the right decision. But there is a strong case to be made here, especially with the company's original mission and non-profit status.
This coup or ousting of Sam, isn’t boosting my faith in this company at all.
Talents exit is a serious concern
Sad. Thanks for posting.
The language in that license agreement is about to become rather important. Numerous lawyers are going to wake up stuck to their sheets tomorrow morning. Not just IP attorneys, but defamation specialists.
Seriously, if a ceo needs to stoop to manipulation to get things done then they’ve lost control of the entire company and manipulation is forestalling the inevitable. Sure, a bad ceo can create large returns for investors but people who hold such myopic views of success navigate life with their training wheels still on.
This thread was detached. I was referring to https://twitter.com/gdb/status/1725736242137182594
Presumably, Microsoft’s lawyers did that before they bought into the novel corporate structure. I mean, its Microsoft, not Elon Musk. Due diligence is a thing.
There’s also the realpolitik of the situation - MSFT owns and controls the compute, in the worst case they can take their ball and go play somewhere else
The announcement affected Microsoft's stock price today.
[1] https://twitter.com/karaswisher/status/1725718391548207246
lol
Which might be bad for my personal short-term employment prospects, so here’s hoping that doesn’t happen. We can ride this BS to the next speculation bubble if we believe in ourselves. And don’t look down.
Maybe he had something to do with it? Maybe, just maybe, it didn't just randomly happened to him.
People shouldn't get board roles based on what degree they have, but based on how well they can do the job.
Helen Toner is famous among the AI safety community for being one of the main people working to halt and reverse any sort of "AI arms race" between the US & China. The recent successes in this regard at the UK AI Safety Summit and the Biden/Xi talks are due in large part to her advocacy.
She is well-connected with Pentagon leaders, who trust her input. She also is one of the hardest-working people among the West's analysts in her efforts to understand and connect with the Chinese side, as she uprooted her life to literally live in Beijing at one point in order to meet with people in the budding Chinese AI Safety community.
Here's an example of her work: AI safeguards: Views inside and outside China (Book chapter) https://www.taylorfrancis.com/chapters/edit/10.4324/97810032...
She's also co-authored several of the most famous "survey" papers which give an overview of AI safety methods: https://scholar.google.com/scholar?hl=en&as_sdt=0%2C5&q=%22h...
She's at roughly the same level of eminence as Dr. Eric Horvitz (Microsoft's Chief Scientific Officer), who has similar goals as her, and who is an advisor to Biden. Comparing the two, Horvitz is more well-connected but Toner is more prolific, and overall they have roughly equal impact.
And that person's tweet chain reeks of speculative elitism.
They discount the trust-based permission society grants Open AI because of its affiliation with the YC and Silicon Valley model just for LLMs.