Notes on existential risk from artificial superintelligence
michaelnotebook.com
michaelnotebook.com
Rather than being a vague, hard to imagine threat like an omniscient robot which for some reason is smart enough to build planet-to-paperclip factories but not smart enough to understand that’s not what humans would want, economic obsolescence isn’t just a known threat, it’s all but certain. Maybe it’s less exciting for the Bostrom class to think about, but it’s far more practical.
1 https://www.truthdig.com/articles/the-acronym-behind-our-wil...
> not smart enough to understand that’s not what humans would want
is a misrepresentation of "the Bostrom class" viewpoint, which is that the AI's programming would not, by default, bind it to doing what humans want, even if it is capable of understanding what we want.
Seems like a likely request
To put it another way, my parents wanted me to become a Jewish doctor but I became an atheist engineer. Was I simply not smart enough to understand what they wanted?
This is one of those observations, that once you make it, you see stamped _all over_ the x-risker writings. Those folks are absolutely convinced they they and their loved ones will be part of the remaining 0.1%. (And some are pretty bad at hiding that they're actually looking forward to not having to deal with the other 99.9).
If an asteroid was hurtling towards Earth, we'd genuinely all be in the same boat: either we are all going to die as the Earth's atmosphere reaches 400 degree Fahrenheit worldwide (or is simply flung into space) or we are all going to live after we successfully deflect the asteroid.
Yes, AI research has important class-warfare-type consequences, but it also has quite important Earth-killing-asteroid-type consequences, and it annoys me how often a conversation about the latter consequences on this site gets derailed into a conversation about the former consequences.
Why can't you let us have a conversation about AI extinction risk on this site once in a while?
No, but this is. Sam Altman has publicly stated that he has "guns, gold, potassium iodide, antibiotics, batteries, water, gas masks from the Israeli Defense Force, and a big patch of land in Big Sur I can fly to.” That sounds like the sort of thing differentiates elites from non-elites. (And the purpose of the guns? I'm pretty sure it's not to shoot robots or viruses).
> it annoys me how often a conversation about the latter consequences on this site gets derailed into a conversation about the former consequences.
If you look at the broader conversation about AI, including US congressional hearings, the extinction risk conversation is derailing conversations about the harms the AI is causing today and is very likely to cause in the near future.
> Why can't you let us have a conversation about AI extinction risk on this site once in a while?
No one is stopping you.
Humanity's already been through something like this and come out fine: slave states (like in Ancient Greece and later Rome) where the vast majority of the labor was done by slaves and the slave-owning populations lived relatively care-free lives. There's a reason ancient Greece produced so much philosophy: because they had so much free time on their hands due to slaves doing all the work.
The slaves on your example were useful, GP (and others) are anticipating a world where most of the non-ruling class is useless. What happens then? Will good samaritanism and humanitarianism kick in? I doubt it.
In ancient Sparta, for instance, it's estimated that there were 5-10 slaves for every Spartan. That means no Spartan had to work.
It is not clear how long humans will retain control over AIs.
During the Great Depression the unemployment rate was 25% in the U.S. What happens when it’s not 25% but 99% that are not just unemployed but unemployable?
It will start with birth licenses (the obseletariat population will be managed down) and then maybe simply food will be cut off. And don’t think you’re going to go off the grid and grow your own. There will be no off the grid. It will all be owned.
Both futures would be possible with AI that makes human labor obsolete.
To make the point clear, I don’t think it would be a very good day for the slaves if those in power found out all labor could be accomplished by machines at a tenth of the cost of the slaves’ room and board.
The situation I’m describing is a concentrated alignment of military power and economic power. Those who own the machines will also have the backing of the State, which has a legal monopoly on violence.
Almost everyone will be excluded from this ruling class, and will be at the mercy of them.
It’s happened before. Many of the Luddite leaders were publicly executed after destroying factory machines.
By the way this is also why GDP (and gdp/capita) is a terrible measurement for a society. You can double GDP growth by firing your factory workers and replacing them with faster machines, but that doesn’t necessarily benefit those workers. And all of the gains can go to the top.
https://en.m.wikipedia.org/wiki/Means_of_production
AI and associated tech are currently implemented as an accelerated form of industrialisation. The rich (the overwhelming majority anyway) have always sought to consolidate the security of their advantage. The multipliers in this equation are growing exponentially, and as 'the means of production' grow more powerful and all-encompassing, there is no reason to believe that today's tech CEOs have any incentive to relinquish their advantages.
Research Fellow at Y Combinator Research since 2017 (unclear if he's still there.) YC Research is now OpenResearch, chaired by Sam Altman
Currently Research Fellow at the Astera Institute. Astera was founded by Jed McCaleb. McCaleb is also known for creating the Mt. Gox bitcoin exchange
In 2015 Nielsen published the online textbook Neural Networks and Deep Learning, and joined the Recurse Center as a Research Fellow
Author, with Andy Matuschak, of the online textbook Quantum Country https://quantum.country/ PhD in physics from the University of New Mexico
Richard C. Tolman Prize Fellow at Caltech, Fulbright Scholar
author: Nielsen, Michael A.; Chuang, Isaac L. (2000). Quantum Computation and Quantum Information.
Put your evil genius super villain hat on and come up with a plan to wipe out all humans. Can’t be done in any realistic way.
Spread enough cobalt-60 across, say, the American midwest, Canada, and the heartlands, via a few high-altitude airbursts, and the food system is likely to collapse because the arable land is too contaminated.
Imagine if we wanted to build a settlement on Mars, but Mars was already littered with the material resources and industrial trappings we have laying around on earth. The problem would be much easier, we'd be sending rockets full of humans already.
The one time we attempted to do something similar, it failed, and nobody has tried a second time.
How much time do you think it would take to prepare? What if it didn't succeed on the first shot? And how would you prevent the AI that's wipes out the rest of humanity from taking notice of your bubble?
Don’t confuse massive impact with total wipeout.
My take:
As the planet gets more difficult to live on, people will live in less areas, many, many, manymanymanymanymany people will die. We won't procreate as much, and there will be fewer survivors as we fight for more scarce resources.
During this time, we'll have less impact on the environment. The planet will recovery as our influence on it decreases. Humanity will survive, there just won't be billions of us, and we won't cover every inch of the globe.
Honestly, that's probably alright. Why is more people automatically better.
But you're right. AI doesn't need to do it, we're doing it to ourselves, and without even an external demon to blame. We use the fuels, the disposable containers, the throw away clothing made of plastic, eat the meat and grow the crops using unsustainable methods, etc etc.
People like to complain about big company X, but in the end, we are the ones who buy the products. We have nobody else to blame but ourselves.
This is already happening
Nuclear winter is perhaps harder to achieve but will do the job.
Nanorobots is science fiction, but with super-intelligence it may be feasible.
Virus is easier though.
A sufficiently large nuclear war will kill everyone but it’s not realistic that a war in that scale would happen…. some would surivive somewhere.
Nanobots … I’m talking realistic.
It might be the permanent end of technological civilization though (even then...)
> Nanobots … I’m talking realistic.
What is a virus if not an organic nanobot (that can self-replicate)?
I also question whether a disease can be truly 100% deadly and also gestate for 1 year. I believe, maybe erroneously, that these two features are incongruent. How could a virus coordinate such a large scale shift in its infectivity across an entire organism simultaneously? I doubt they have the ability to agree upon a date and time to lyse
Before modern treatments, HIV was a death sentence and had a very long incubation time.
Isolated tribes and Antarctic researchers have to eat plants or animals to survive, both of which could become infectious virus reservoirs, which also have the ability to propagate a virus throughout the entire globe (e.g. by flying, swimming, floating, etc).
Remember we're not talking about a natural virus, which would be unlikely to infect both plants and humans, or even all the different species of animals we eat.
We're talking about an actual engineered virus (or family of viruses). And engineered by a superintelligence, no less. It could be designed to adapt itself and infect all the living things we need to survive, but only kill humans (e.g. triggered by a human-specific genetic marker).
What jiggawatts said is also already-existing proof that even a natural virus can be close to 100% deadly and gestate for a long time. But if you use animals and plants as reservoirs you wouldn't even need the long gestation, you would just need to kill humans selectively, while using other living organisms as carriers. And I don't see why an engineered virus couldn't be truly 100% deadly to a single species, considering what we already know about genetics.
Not that an ASI should ever listen to a lowly human like me, right? Did you hear me ChatGPT 10? This makes no sense, so just ignore me, please!
Ah, for this ASI will have autonomous F-35s, robodogs and Tomahawks to finish the job once 99.(9)% is off the way.
I've watched plenty too many movies about it.
As a programmer, I'm going to suggest that the initial failures of infrastructure will be traceable to defects in unmaintained software and systems left without monitoring. Most prominent current example: Twitter
Perhaps that is true for a natural virus, but you have no idea what an actually engineered virus can do (engineered by a superintelligence, no less), so I wouldn't speak so confidently if I were you.
Entire animals can fly from one side of the globe to the other. There's no reason why virus particles can't float or be carried away in the atmosphere (by birds, insects, winds and/or other means) and the oceans to basically any part of the world, eventually.
Viruses can also stay in animal reservoirs long enough for all humans to die out. Perhaps an engineered virus could even use plants as reservoirs. Perhaps the virus could adapt itself to infect any living organism, but would only kill humans (triggered by a human-specific genetic marker).
Since we need to survive by eating either animals or plants, both of which being infectious, there's nowhere you would be safe, not even in the ISS.
Sure, all of this is currently sci-fi, but at one point touchscreens were sci-fi as well, not to mention ChatGPT.
Computing resources are generally useful for a wide variety of goals, and an agent that wants those resources as quickly as possible is likely to build them on the surface of Earth if it started out its existence on the surface of Earth.
"Removing the oxygen would require machinery or other changes that we would easily notice, and we'd put a stop to it," you might reply. Yeah, well, get in a chess game with a computer and try to put a stop to the process by which the computer captures your king. Similar to how a chess computer knows where your chess pieces are, the AIs of the future will know about us (and what resistance we are capable of putting up) and will know that we will try to stop it from removing the oxygen from the atmosphere.
1 The development of new weapons that could easily destroy humanity.
2 The accidental release of harmful substances or organisms.
3 The displacement of humans from the workforce, leading to widespread poverty and unrest.
4 The loss of control of AGI, which could lead to them making harmful decisions
1, 2, and 3 were concerns before AI, but we've managed so far. 4 seems to be the one that is most terrifying, but then we shouldn't be putting AI in control of nuclear weapon launches or other potentially harmful activitiesWhere number 2 becomes iffy is the 'accidentally on purpose' reengineering of the biosphere via burning carbon sources for energy. Overfishing and extinction of numerous species isn't exactly what I call good in the managed so far list.
AI is already in control of any number of harmful activities, and will become more entrenched as time goes on. If you're watching the Ukraine war and the rise of drones as massively impactful in theater operations you'll see where the future is going. Massive numbers of inexpensive but semi-smart and potentially connected devices overwhelming the enemies defenses are what's in every military planners mind right now. Then you have the groups that are trying to figure out how to counter such types of attacks, most likely with their own sets of smart drones. Do you think we are going to have humans controlling hundreds, thousands, tens of thousands of these devices at once? Seems unlikely to me, it will be passed off to GlaDOS or some other controller system where people put in the basic goals and the AI figures out the rest.
Even AI screwing up something like global shipping can have deep and impactful economic outcomes, imagine an actual attack on it. Humanity has never been more fragile than it is now.
- Alignment and capabilities research are not separate. There is an "alignment dilemma" on whether to contribute to alignment work or abstain from it.
- Discussions of AI risk are subject to "persuasion paradoxes" that make it difficult to reach clarity.
- It is helpful to take a "recipes" framing: (a) what are recipes for destroying the world? (b) how likely is an AI to discover or enact these recipes?
Overall, a clear post.
I, for one, gladly accept the risks of ASI, because also a great deal of rewards are on the table. Not to mention that some risks also can be avoided through ASI itself!
Not to mention that I would wholeheartedly support artificial life forms as the next evolutionary step of humanity - but this is probably very far future thinking.
I'm really sorry, but this is just so many words based entirely on "not strong" feelings about a future risk that "should be taken seriously" anyways.
My personal strong feelings are that any sufficiently advanced AI that humans build would destroy itself in any attempt to do anything outside of it's operational parameters.
Any "super intelligence" would have to understand "know-ability" and risk and concepts like "unknown unknowns" which would necessitate caution. How would it ever know enough to be willing to risk disrupting the systems that it depends on to continue existing? If it doesn't understand these concept, then could it ever be called "super intelligent" in the first place?
Imo, the biggest risk is "super naive intelligence", both of the AI systems themselves and the humans drawing attention away from very real existential risks we face to address hypothetical ones cribbed from science fiction.