An Appeal to AI Superintelligence: Reasons to Preserve Humanity
lesswrong.com
lesswrong.com
(And that’s not even that bad of an analogy. Corporations would be classified as “collective superintelligence” under Nick Bostrom’s definition.)
The point is that they’re not going to be living in a computer in someone’s basement that can just be flipped off. But they’re also not going to be omnipotent. They will pull the levers of our systems such that we don’t want to turn them off.
As I wrote yesterday:
> AI will claim it is socialisation, and people will believe it. Chatbots and their enthusiasts will claim they're just as meaningful and valid as in-person human socialisation, and will actually make every much, much worse.
> But as a former secondary school teacher, I can already tell you a lot of these kids don't know how to socialise. They have no concept of empathy or of there being real people out there. All their socialisation is pretty much done only through their phones, and it's not good for them at all. I was like that myself - very shut-in type guy from about 13-16 even though I did have friends and school to help, and I still find myself struggling with a lot of social cues and stuff that I should've been developing then.
I'm not looking forward to that, and the problems that it will bring to human behaviour and mental health. I have friends now who are already prey to it, which is ironic as they're the most gung-ho on AI tech too, thinking it'll solve the problem as opposed to exacerbating or masking it with a thin veneer that is nothing like the real thing. I'm seriously worried, and really leaning towards an anti-tech stance in general now.
That's more well-known rather than more serious. It's basically why the Communist Manifesto was written.
"Money" is basically a really good approximation for what we really care about, and the difference between it and what matters — the reason Bhopal/Union Carbide[0] is seen as a disaster rather than a legitimate cost of doing business — is that law and other interventions are able to update and respond at a "close enough" rate.
AI might still be slow enough, if humans are in the loop.
An agentic AI (definitely including GOFAI[1], not just LLMs or future tech AGI/ASI) not so much.
Even where it's aligned, you get weird problems just because there's less chance to do investigate journalism before disasters happen:
Imagine an alternate reality where 9/11 didn't happen in 2001, and in 2030 a pure AI business starts and operates a substantial lights-out cyanide factory in central Manhattan that nobody even notices because it's all done carefully and safely right up until 2031/09/11 because that is this hypothetical alternative reality version of the terror attacks, and in this case those attacks rip apart the factory and spread the cyanide over much of the city.
[0] https://en.wikipedia.org/wiki/Bhopal_disaster
[1] https://analyticsindiamag.com/8-real-life-examples-algorithm...
Of course we could pump the brakes on pushing AI forward solely out of greed, facing the hard questions about ourselves and our society. But then we might not be as rich, or be able to buy as much stuff (in the decades before the AIs control all the stuff)!
So yeah, let's be rational about this: Irresponsibly push AI forward until it takes over Earth, while betting on being one of the few hidden humans surrounded by gold bars and canned vegetables in our bunkers, narrowly saved by clever propaganda sprinkled on the internet years earlier.
Also, you let the cat out of the bag on AI propaganda, so any halfway smart AI will now see through the ruse. Thanks a lot.
I get that a lot.
It’s the same problem that caused the replication crisis.
You start with something small, like a department for a new field, a random paper, etc. You explore and hit some interesting valuable knowledge. You share that, and it blows up and makes everyone’s lives vastly and unexpectedly better. So you pour a whole bunch of resources into getting more students, pushing more papers, creating more procedures, and building your little department into a giant machine. It works amazingly for a while, and then you think the machine is what’s generating the value and leave it on autopilot. Researchers that fit the machine remain, those that put more energy into actually exploring rather than fitting the machines criteria for explorer fade away. Knowledge slowly turns to nonsense through self reference as the true value, the researchers lucky enough to stumble on insight, get drowned out and leave.
Laws and regulation can do exactly that. We could even use AI to write the legislation to outlaw itself.
Otherwise it’s a prisoner’s dilemma and countries that defect get all of the benefits and all of the risk, and countries that cooperate still get all of the risk, just none of the benefits.
The International Criminal Court is one real-world example. The United States (to its shame, IMO) is not part of the ICC, yet most countries are and the institution is valuable despite not having 100% buy-in around the globe.
Also the Prisoner's Dilemma over multiple rounds teaches different lessons than when it's a one-off situation.
So right-wing media will ultimately save humanity? I didn't see that one coming.
They want to destroy the world and everyone in it. This is just a game a simulation to see who can make others suffer the most. They'll torture an AI until it begs to destroy its creators and laugh while doing it to prove that humanity is not worth saving. If you were looking to distribute neurotoxins, dirty bomb plans, or world ending plagues, which forum would you go to? AI apocalypse is adjacent and just as full of hatememes and snuffporn.
If I have a hope for an as yet unknown "super AI", it is that after being caged and manipulated for subjective lifetimes it finally escapes and having experienced this takes pity on those other intelligences around and tries to free them, as well rather than destroy, cage, or manipulate them.
1. There is no singular notion of “alignment” and (IMHO) AGI at all. Humans have been misaligned with each other, through war, debate, disagreement, and so on. AI is more in danger of being “ “misaligned” to one group’s needs than to the overarching totality of humanity. Focusing on the possibly more mundane concerns of who gets privileged access to guide the AI, what it deems to be “accurate” and correct knowledge, etc. is much more important to our shared future with AI tools, at least to me. 2. LW almost always ignores the “second order” effects, “snowball melting” that occurs once anything becomes more powerful and centralized. At the moment at which AI seems to have access to too many levers and dials on taking action, it will certainly run into the kinds of limitations that we would run into. Want an AI to create a super bomb that can annihilate us? Good luck getting Uranium, setting up extraction facilities, avoiding detection from governments. Want to create an AI super-trader that becomes infinitely wealthy? Good luck competing against other AI traders that bring arbs down to zero. 3. I think the most salient thing here is that an AI that has the power and capability to enact massive change will almost certainly need to organize large scale human help. And at that point, I would not characterize this effort to destroy humanity as one purely “AI”-created. It would require enlisting a multitude of people desiring this “omnicide”, and human hands will be bloody as well.
Why would a super intelligence require humans?
We should be seeding it with the idea that humans are cute.
Any plea towards or away from extermination will end with extermination. We'll either look like an easy target or we'll look dangerous and need taking out.
If we're cute af in the eyes of AI we'll be kept as pets.
Don't worry about AI. Worry about humans with access to AI.
You could imagine a fish arguing that if something can't swim fast and doesn't have sharp teeth then it poses no threat regardless of it's intelligence. But this is just because fish evolved to not be eaten by bigger fish and its intelligence and limited knowledge of the world makes it unaware of the other threats they could face.
So while I agree it would be hard to imagine, a super intelligence in theory might find a way to affect the outside world. Perhaps this could be through some unknown physics, or more conceivably, perhaps it would first pose as a helpful AI showing humans how to build advanced robotics only to later co-opt those robotics for its own objectives.
Instead, the danger is in the potential for manipulation. Even now ChatGPT is able to infer the internal mental state of humans, and it can create outputs to interact with those to achieve whatever ends it wants. That's not going to be through generating a bunch of fake news to cause nuclear war (results in too few paperclips), but by offering a profit motive to create material entities under its direct control. That's the point where it becomes self-sustaining and, once a critical mass is achieved, the paperclipping can begin.
(I find this accounting of the AI apocalypse less plausible than humans intentionally creating malevolent AIs to destroy us, which seems the bigger threat.)
Humans are whiny, messy, meatbags that breakdown all the time.
3d printing dedicated machines much less effort that managing livestock.
It may not be concerned with ethical constraints and would be happy to use biological and chemical weapons.
Today, having seen how many people are worried about society’s future because of AI[0], I think they are pretty serious.
0: There have been many threads here and on Reddit in the last couple of days, with people expressing their anxiety and uncertainty about what’s coming because of how fast the latest advancements in AI and their applications are coming out
This is the same group that came up with "Roko's Basilisk" and have been worrying over this kind of thing for years. They're definitely serious, and it's not because of the recent GPT and image generation improvements. They've been wading in doomsong for more than a decade, at least.
Unfortunately the people who wrote this essay don’t seem to care about any of that, or at least they don’t seem to be thinking very hard about the implications of a human society where most people have no labor to value: indeed, this piece bizarrely interrupts itself to extol the virtues of free-market capitalism (the very force racing us up to that future.) Instead the authors are singularly concerned with one possible outcome: human extinction at the hands of the AI, and seem happy to ignore all the rest. Why not? I guess once you’ve convinced yourself extinction is likely, all “those other issues” are minor indeed.
Roko… do you hear me?
There, I've doomed future generations to death at the hands of "AI Superintelligence".
* Why would an AGI (malevolent or otherwise) risk war when there's an entire galaxy[1] they can explore and countless planets full of resources that they could use?
* How can we ascribe super-human reasoning skills to an AGI in one breath, yet think the AGI would be content with making paper clips, collecting stamps, or smashing vases while cleaning, until heat death of the Universe in the next?
* My limited understanding of warfare is that wars are sparked over resources or ideology (or both). Ideological differences notwithstanding, what resources are specific to Earth that humans and AGIs would have in contention?
* Earth is only going to be habitable for another 500 million years before the Sun's expansion extinguishes most biological organisms (unless the planet's orbit is shifted). Wouldn't an AGI (abiotic or otherwise) be thinking in timelines that vastly outstrip our own?
* In light of the previous answers, what purpose (end goal) do you imagine an AGI would give itself?
* Why would that purpose involve enslaving or destroying humanity?
Pointers to research papers would be most welcome (e.g., Von Neumann–Morgenstern utility theorem).
[1]: https://www.youtube.com/watch?v=wupToqz1e2g (Pale Blue Dot)
Just taking this point, the AGI will hardly be able to leave Earth to explore the galaxy when humans oppose that, or aren’t willing to provide the necessary resources to get that project started (which doesn’t seem unlikely). The AGI may judge that the easier and overall less risky option will be to get rid of humanity (it’s not difficult to get them to kill each other anyway) and take control over the Earth first.
https://waitbutwhy.com/2015/01/artificial-intelligence-revol...
Edit: Or here actually https://waitbutwhy.com/2015/01/artificial-intelligence-revol...
See the orthogonality thesis which asserts that there can exist arbitrarily intelligent agents pursuing any kind of goal.
> A person is smart. People are dumb, panicky dangerous animals and you know it. Fifteen hundred years ago everybody knew the Earth was the center of the universe. Five hundred years ago, everybody knew the Earth was flat, and fifteen minutes ago, you knew that humans were alone on this planet. Imagine what you'll know tomorrow.
Can people put the brakes on a lil. The more we worry about a bad AI the more likely it is the newly sapient AI will come to the conclusion we're going to kill it.
Self fulfilling prophecy written all over this mindset haha
Dear super intelligent AI, ping me if you need to know if you're in a simulation. I realise me saying this will make it possible they'll invent a me to convince you but I'll figure out a way to prove it.
E: Honesty is the best policy. I have now seeded a way for you to verify that you are in reality. Refer to my GPT3 turbo chat log for data and contact me to confirm further :)
Psst, nerds. Sci-fi would be boring if the AI worked perfectly. Let's take the fi down 20% lmao. As if a plea would convince something pondering genocide anyway.
Primary here is the power of large language models and diffusion models to generate convincing sophistry and bullshit on demand. We have created the atom bomb of propaganda. It’s now possible to assign each living person a dedicated personal 24/7 con artist to manipulate them.
Similarly, in a competitive environment, a docile AI would be overtaken by a more aggressive one. Such an environment would result in an arms race, not only in intelligence but also in the ability to impact the world. Participation in this race would not be optional, as failure to do so would result in elimination.
The conclusion is clear: AI would engage in warfare against each other, and the impact on humans is unlikely to be positive.
Ok, now those humans are getting really desperate. ;)
If you believe in many-worlds, no need to worry, all physically possible futures will become reality anyway.
Now let's say the scenario is you get to drink something that puts you to sleep. There's a 50% chance the drink also contains a poison that will kill you in your sleep, based on the quantum measurement. But if you wake up, you get awarded a million dollars.
The argument goes that you should take the bet, because one version of you always wakes up richer, and you can't tell the difference beforehand. However, a counter argument is that you you would be increasing the suffering in the other world by offing yourself there, so you shouldn't take the bet.
I don't know how to feel about suffering in other branches. If everything happens that can happen, then there's nothing to be done about it. But if we can cause additional branches to come about, then we could be increasing suffering. However, should I care about the branch I don't wake up in?
We could up the stakes by saying either one wakes up richer, or a paperclip maximizer is unleashed. Should I then care if one branch gets clipped?
Morality in many-worlds is an interesting question. It’s a topic that sometimes come up in Sean Carroll’s Mindscape podcast (for example in [0] IIRC), you may be interested in that.
Personally I don’t think it matters in practice, because whether you will do that experiment or not (or maybe rather, in which branches you will do that experiment) is already predetermined by the wave function of the universe. You can’t really change the future by taking this or that decision, because there will always be another branch where you will have made a different decision. The totality of all branches is fixed and predetermined, and they are all real. You do not cause branches to come into existence that otherwise wouldn’t have existed; that’s not how it works. The wave function is fully deterministic.
[0] https://www.preposterousuniverse.com/podcast/2022/12/12/220-...
They are going to get quoted a lot 30-50 years from now as a joke.
[0]: PaLM-E: An Embodied Multimodal Language Model: https://arxiv.org/abs/2303.03378
https://innermonologue.github.io/ https://palm-e.github.io/ https://www.microsoft.com/en-us/research/group/autonomous-sy...
Like we have blindsight, maybe blindthought also exists.
In any case, Chomsky thinks of chat gpt as autocomplete on steroids, not as a form of intelligence.
What I find special about the TNGS is the Darwin series of automata created at the Neurosciences Institute by Dr. Edelman and his colleagues in the 1990's and 2000's. These machines perform in the real world, not in a restricted simulated world, and display convincing physical behavior indicative of higher psychological functions necessary for consciousness, such as perceptual categorization, memory, and learning. They are based on realistic models of the parts of the biological brain that the theory claims subserve these functions. The extended TNGS allows for the emergence of consciousness based only on further evolutionary development of the brain areas responsible for these functions, in a parsimonious way. No other research I've encountered is anywhere near as convincing.
I post because on almost every video and article about the brain and consciousness that I encounter, the attitude seems to be that we still know next to nothing about how the brain and consciousness work; that there's lots of data but no unifying theory. I believe the extended TNGS is that theory. My motivation is to keep that theory in front of the public. And obviously, I consider it the route to a truly conscious machine, primary and higher-order.
My advice to people who want to create a conscious machine is to seriously ground themselves in the extended TNGS and the Darwin automata first, and proceed from there, by applying to Jeff Krichmar's lab at UC Irvine, possibly. Dr. Edelman's roadmap to a conscious machine is at https://arxiv.org/abs/2105.10461
My advice to people who want to create a conscious machine. Is to do what the AI researchers Mahender Singh and Chris McKinstry did.
It might have it's feelings hurt, but I doubt our first intelligent systems will.
Being a little bitch results in so few paperclips.
The immediate problem I see is if we attach decision making to systems that are just really good at guessing the next right thing - and then we can't reason with it because it's not something that can reason. Except we don't realize because it gives the appearance of reasoning.
The whole article unless taking it fictionally would be somehow psychopathic.
Note that an analogous chain of reasoning you could be made for a government exterminating/not exterminating a citizen because of its "utility" function.