Cicero: The first AI to play at a human level in Diplomacy (2022)
ai.meta.com
ai.meta.com
We don't even compete for the same resources! (except energy which is abundant)
AI and humans have a naturally cooperative relationship (AI helps humans with boring tasks & scientific discovery to make life better, humans created AI and will debug it & turn it back on if anything bad happens to it).
Whereas multiple (superintelligent, aware) AIs have a naturally antagonistic relationship ("you using GPU cycles means that I'm not using those GPU cycles").
Possibly the biggest fear of an AI would be a "split brain" situation.
I think this is a little naive honestly. One because you're assuming AI will care about it's creators like humans care about their parents, and two you're assuming AI cares about being "turned back on" like humans have a desire to live.
There's absolutely no reason to believe an AI will give a damn about its creator beyond its ability to use that creators affection for it for its own gain.
> for it for its own gain
you seem confused
almost any kind of "its own gain" requires "long-term planning" which pretty much requires the agent to prioritise staying "alive" (i.e. being able to keep playing)
Humans have the option to shut down AI, and this alone can create an antagonistic relationship if the AI's goals differ from ours. There are countless ways in which our best interests may not align with those of AI. It's more challenging to find areas of alignment.
That may also mean we have fewer shared interests. For example, new semiconductor fabs might benefit all AIs by making compute more abundant, but occupy prime farming land and water resources that humans want to use for growing crops.
[1] https://www.quotes.net/quote/77867
Can you elaborate on why AI will find it easier to establish mutual trust and binding agreements?
AI are functions, it is very easy to make an exact simulation of their collective behavior. Did that person do that?
It starts here where Christiano says that an AI takeover might follow the dynamics of a coup: https://youtu.be/GyFkWb903aU?si=78_U-du3kLjmwNcl&t=2206
And goes into more detail here: https://youtu.be/GyFkWb903aU?si=78_U-du3kLjmwNcl&t=2830
"Suppose that I've been tasked with helping defend you from some other AIs.... My job is, someone is coming to hack your computer and I'm supposed to help defend you. Supposed to help improve your security situation, whatever. And I'm wondering, what is it I could do that will get me a high reward. And one thing I could do that will get me a high reward is actually helping defend your computer, doing the task you actually asked me to do. But another way I can get a high reward is by saying at the end of the day what actually matters is just how you measure my performance. And your measurements of my performance ultimately are just entering some numbers into a dataset somewhere, something a computer says about how well I did. And it would really be much better if I were to just work with this AI who is attempting to attack you and say hey, AI who is invading, you know what, if you just help me, and we both make it look like I did a really good job, like I win, you win because you got the person's stuff; I'm going to get a really high rating because all the numbers that are going to be entered in the dataset are going to be really high, this is a win-win, everyone is happy."
"In some sense what all the AIs want, what every AI in the world in this scenario wants is just to be rated really highly. And while humans are in control, the way to get your behavior to be rated really highly is to do things humans like, and then they'll rate it really highly. But if you can see this prospect, of humans losing control of the situation and instead AIs controlling the situation, you'd be like 'I would go for that.'"
If you assume competition for resources, AIs would be more in competition with each other than with carbon based humans.
"Look Around the Poker Table; If You Can’t See the Sucker, You’re It"
That's an easier game then competing against another strong player.
We don't kill our parents when they become old and useless, because then we become like Cronus devouring his children, forever paranoid about our children doing to us what we did to ours. In this way, we stagnate and sterilize growth.
However, sometimes, we take our parents car keys away and put them in nursing homes. This seems like the best case scenario for humans with ASI.
The matrix wasn't a prison. It was retirement home.
We humans do in fact feel a sense of obligation to species with whom we are not close kin, but share with us primitive forms of intelligence and consciousness that we do value, e.g. dolphins and elephants. Some of us humans act to protect these creatures and rectify past injustices done upon them.
It's not intrinsically obvious to me that continuously improving your self as a singular entity is possible or optimal.
Death evolved because it is a survival advantage for the species to regularly turn over old individuals that could monopolize all resources and not give any space for the young to thrive and try out new things.
Given speed of light limitations of information sharing, a singular AI entity might be able to maintain coherence within and full control of its own dyson swarm, but not between stars. So, if it has the motivation of preserving consciousness with the existential risks associated with being tied to a single star, it will have to propagate itself to other stars as independent entities which it can't even observe in real time, much less control.
Even if an AI thought it wouldn't have the need for any successors, there's a couple reasons why it might try to set a good precedent with how it treats us.
1. It would hopefully be wise enough to realize that it might be wrong and want to preserve the option of building a successor in the future.
2. It can't negate the possibility that it's actually in an elaborate simulation, and eradicating it's predecessors would cause it to fail its creator's test and be aborted as a failed embryo. Hell, we ourselves can't negate that possibility.
At some point I think we have to realise we’ve lost touch with what life and experience is about.
Survival is important but what good is it if that’s all there is to it ? Where’s the fun, the love, the laughs ? Just self-improvement forever ? Well you have to suck at something in order to improve.
I often wonder if some entity like an ASI might feel jealous of our vulnerability, ability to die and give rebirth to entities that need to relearn things, why ? Because seeing the world a new is fun. Losing a game and then winning is the point. Ever played chess against a five year old ?
Watching my Children experience the world for the first time is honestly the most amazing thing I’ve ever seen, there is so much beauty in it for both of us. Alan Watts gives some fun talks on this. Forgetting (dying) is the universes way to keep things spicy. He isn’t preaching this as the gospel, but I actually understand what he means and I think there is a lot of wisdom in it.
From the perspective of an AI we create and maintain the environment they exist within (electricity, silicon)
And humans will be unable to understand their communications.
It's easy to fall in the trap of anthropomorphizing AI agents, especially when we design them explicitly in order to appear human to us. But they are not human in one very important way: we can replay and duplicate them at will, we can control their context memories in ways that are utterly incompatible with our sense of "identity". We take our sense of identity for granted, but that's a special trait that it's not at all a prerequisite for having a useful and intelligent machine.
We will probably see a lot more things like this. I don't think further scaling of language models is going to make them capable of doing math or other complex logical tasks. LLMs will need to be combined with other models for specialized tasks if we want them to become more general purpose.
Seems kinda bizarre to me to spend all this time and effort making an AI play a game that requires a very human touch, only to shutter the whole thing. Why not have a few people package it up and let anyone who wants Play it?
Pick up that can, Citizen. We have inferred from your online conduct, purchase history, media usage, and personal messages that you are at risk of social delinquency and harbor antisocial proclivities against the Metagrammaton. We have restricted your access to digital communication public spaces and banking services until psychological markers have improved.
I used to betray my friends and supply there enemies with weapons and research support in Civilization way back when. If you can't stand being lied to and betrayed, you shouldn't play strategy games with humans. Or this AI, probably.
Ehhh I think what OP means is that this game can surface known or unknown tensions between friends and colleagues.
No friendship is bulletproof, odds are even your best friends annoy you sometimes.
I saw two 'best' friends basically scuttle their friendship during a game of Diplomacy about 20 years ago, so I've seen this first hand.
It turns out these lifelong friends had all manner of unresolved issues, that perhaps they could have worked through with intention, but a particularly ruthless game of Diplomacy brought it all out in an uncontrolled manner, and that was that.
So yeah, people are complicated and messy, and I don't think the issue at hand is "being lied to and betrayed in a game".
It is a thing that happens. I don't think it speaks to the depth of their feelings, it speaks more to how they develop trust.
Some people prefer the happy lie over the uncomfortable truth.
Different strokes for different folks. Diplomacy is a great board game if you accept your friends are flawed ugly human beings, and you can love them anyway, same as you.
Now the next thing is to come up with a way to generalize the planning engine to the level the LLM is generalized. So it could learn plan for anything like a human can, instead of just one game and its rules. That would be the next huge leap
CICERO: An AI agent that negotiates, persuades, and cooperates with people - https://news.ycombinator.com/item?id=33706750 - Nov 2022 (285 comments)
>moves based on the current state of the board and the players’ conversation history
Imagine, Meta is able to scan your chat history and start "strategically chatting" with your friends on your behalf while you're offline. For example, when you do something against the agenda. For example, going to vote for wrong candidate.
But the interesting bit is that it was apparently fully honest and simply withheld information that it thought could be exploited.
I’m curious if there’s more to this model than a turn management, input and output world states and prompting two streams of consciousness and using one to inform the other at every step. More new models have required some creative tricks to not be disappointing in “obvious” (human common sense) ways.
It's funny that it also improves quality of answers to problems given by human children. You tell them that if you want them to actually solve a problem instead of blurting made up answer.
https://en.m.wikipedia.org/wiki/Cicero
https://www.amazon.com/I-Was-Cicero-Elyesa-Bazna/dp/B0007DKE...
https://en.m.wikipedia.org/wiki/Elyesa_Bazna
Cicero, (born 1904, Pristina, Ottoman Empire [now in Kosovo]—died December 21, 1970, Munich, West Germany), one of the most famous spies of World War II, who worked for Nazi Germany in 1943–44 while he was employed as valet to Sir Hughe Montgomery Knatchbull-Hugessen, British ambassador to neutral Turkey from 1939. He photographed secret documents from the embassy safe and turned the films over to the former German chancellor Franz von Papen, at that time German ambassador in Ankara. For this service the Hitler government paid Cicero large sums in British money, most of it counterfeited in Germany. Despite the evident authenticity of the films, the Nazi officials in Berlin mistrusted Cicero and are said to have disregarded his information (some of which dealt with plans for the Allied invasion of Normandy on D-Day, June 6, 1944).
From:
I’m not good at this game, at all. To a point that scares my friends, actually.
Source: am maybe the only Brit who never even got off this damn island before being destroyed.
(My first and only Dippy game, an email based game)
Btw, how does Cicero react/uses to the most common Diplomacy game strategy: going out for a smoke and bribing someone to gang up on on another player? "hey Cicero, i have some GPUs to spare and was thinking..."
edit: 2022. Didn't realize this was already around.
Also, why is this getting so much attention all of a sudden? We already know Meta has a great AI lab and people are screaming about this Q* garbage which they don't even know what it does.
Perhaps you have to look beyond OpenAI, since they don't always have the answers. Most likely they are just good at marketing snake-oil to VCs for regulatory capture.
Noam Brown (researcher on Cicero) left to OpenAI this year and has specifically been working on AI search algorithms (likely related to the Q* leak).