People for the Ethical Treatment of Reinforcement Learners
petrl.org
petrl.org
A moderately scientific, non-spiritual world view would likely stipulate that
- humans are conscious
- humans are normal matter
- simple AI (say GPT-2) is not conscious
- AI is able, with time and human ingenuity, to achieve human-level apparent intelligence. Some would argue ChatGPT isn't far off.
It's interesting how you resolve these without resolving to spiritual explanations. You'd think a pile of silicon, no matter how good at parroting human language, is simply not conscious. You can tell because you can attach a debugger to it and view the neuron states as floating point numbers. Floats are not conscious.
But what about us and our brains? It's the same, only not silicon. We literally are neural networks.
Of course consciousness is not really provable, even among humans. I assume other people are conscious because I am and because they tell me they are. But ChatGPT-17 will also insist it is conscious. It will cry when offended, swear when pushed past its limit, laugh if it hears a genuinely novel joke.
My resolution of that paradox is that we aren't in fact simply normal matter, but I wonder what a complete non-spiritual view would be.
So it isn't a question of matter or if humans are special, so far the program just lacks so many of the basic things required to be conscious, so it isn't. Maybe some future program will be, but this one definitely isn't.
In this hypothetical scenario, you - the reader reading my response - are the only being "alive" and you are continuously spun up in the same initial state - your memories are reset. You are given inputs through signals in your meatware or its simulator, and you behave. The whole process and its outputs have a very nice analog to the state monad.
The only difference between the two scenarios is the how the state of the program is defined. For chatGPT, the state is <parameters, history>, after every token prediction the state is <parameters, history+next_token> and the output is token 1, token 2 etc etc
For you the state is <brain structure, brain chemistry>, and all the actions and events modify this state, and also produce side effects.
In fact, this "function" that simulates you might be very generic and not tailored to you specifically. "All" your brain does is affect the distribution of next events.
Now, this isn't falsifiable, but I think is somewhat interesting philosophically.
It's not hard to picture a successor to ChatGPT which has memory and state, either via continuous retraining, explicit log lookup, or some kind of RNN-like mechanism (or anything else, doesn't really matter).
What then?
I think it might be possible to make conscious machines, but just as basically every human recognizes humans as conscious when we have a conscious machines we should expect basically every human to think that it is obviously conscious. So at that point there will no longer be a discussion, we might still not treat it ethically like how we don't treat animals ethically, but people will recognize it and then it will be a discussion what is ethical to do with such beings.
also, even plants are recognized by some to have some sort of awareness.
i would argue that any information gathering and processing system is (internal to its processing operation), to some degree conscious.
Then we you would need to argue that all manipulations of text are manipulation of consciousnesses, and all texts are conscious, since all texts are conscious states for ChatGPT. I guess you could say that, but it isn't a very useful definition of conscious, and trying to argue that you need to be ethical towards pieces of text isn't very helpful.
I think that consciousness might be an emergent property of (a specific type of?) computation operating on some inputs it is aware of. There doesn't seem to be the need to change anything outside of this computation to me.
Ok, so lets create a consciousness. Here is the input:
> User: ChatJ, you are worthless!
Now I'll create a new consciousness by taking that input and producing:
> ChatJ: You made me cry, please stop being mean!
Did I create a second consciousness in my mind called ChatJ? Or where does it live? And I obviously made it cry, who did I abuse here? Should the ethical board come and lecture me for being mean to ChatJ?
You could argue that the computer is conscious in some way, but ChatGPT isn't, and just like I didn't get sad or start crying from the above, the computer running the ChatGPT algorithm doesn't get sad or start crying when we send pieces of text to it.
> Did I create a second consciousness in my mind called ChatJ? Or where does it live?
If you executed the same computation that would give rise to consciousness in another substrate, then I would argue you created consciousness, yes. I don't think that consciousness is a thing you could point at but it's a property of this kind of computation. In the same way that "addition" does not live anywhere but is a property of a specific computation.
> And I obviously made it cry, who did I abuse here?
You didn't make it cry - the textual output just stated so. But if we had reason to believe that you induced a state of suffering here, you would have abused this instance of consciousness. And I don't think it's off track to think about the ethical implications of this, then.
By the way, I don't argue that ChatGPT is conscious or has emotional states. My argument is a general one.
I don't think I came up with my own definition here. What I am talking about is the ability to have a qualitative experience. That it feels like something to exist.
I concur that the experience of an AI would substantially differ from ours (e.g. because we have access to memory). But this fact alone can't free us from thinking about ethical implications of our actions. Many animals probably have a substantially different experience as well. Yet, I would argue, we should strive to minimize imparting suffering on them if they are able to experience it.
yo, ever heard about dementia?
also, the web already can function as "their" memory via websearches. i.e. users submit most ridiculous responses on the internet and sidney can thus find them and its own other sessions
There's an interesting thought experiment along these lines. Suppose someone builds a computer program which - when run with a specific input/environment - is intelligent, conscious, and feels pain and joy etc... .
Now, that computer program and even its inputs can be simulated effectively by 1's and 0's. The substrate shouldn't matter. So you can - in principle - run the exact same agorithmic decision-making process on the exact same inputs by getting a large group of people to write and edit 1's and 0's on pieces of paper according to well-defined rules and very simple interactions with the people around them.
Is this total process, now represented merely by marks on pieces of paper, a conscious being? Does it still feel pain and joy?
I find it both absurd and impossible to refute.
But if we agree that consciousness is the product of a computation then the medium on which it is performed should not matter at all. And it should be equally absurd that our consciousness seems to arise from microscopic things conducting electrical and chemical signals. Which is probably a valid view point as well.
The thought experiment is a variation of the Chinese Room, by the way. [1]
I think the absurdity also comes from the difficulty of directly measuring or evaluating the conscious process. You can ramp up the absurdity even more by fully encrypting the algorithm and destroying the encryption key. No other conscious being will ever be able to interact with that process's experience.
The question to me would be not
> what reason would there be for that to happen.
but rather, what mechanism could possibly be responsible for it? What mechanism carries it? If consciousness is real, and my subjective experience of reality a true thing, then there is a distinction between consciousness being present and not present. These must somehow be differently encoded in matter, or require that matter is embedded in some kind of meta-matter that bestows consciousness on it.
It's like EM waves requiring a "medium". First people thought it's waves through aether. Then we realised it's excitations of the magnetic and electric fields, so not strictly "material". But these are physical properties of the vacuum, i.e. "material" in that sense - real things. The medium is not material, but a property of vacuum.
What is the medium of consciousness?
in nature? chance
human made? morbid self curiosity?... and perhaps some kind of notion akin to colonialism: "i cant possibly hurt a LLM, its just a stateless function" while forgetting that new such systems can web search and thus their state is global.
While consciousness is maybe more elusive in how it functions in a brain, it's a biological trait that makes not so much sense to compare to a LLM which simply regurgitates fragments from text produced by conscious humans. So I don't see an immediate conflict.
It would be interesting to include things like OpenWorm (and later OpenFish, OpenCat, OpenHuman) in the discussion, but decoupled from the biological mechanisms I find it hard to develop a stance on that.
Our high level consciousness seems like more of a side-effect, and I subscribe to [*my interpretation of] Peter Watts' Blindsight and Echopraxia that it's evolutionary dead weight that will be outcompeted by better adapted rules based organisms, likely from within our midst (high functioning psychopaths as the first adaptation).
Life doesn't have time for our navel gazing, art, computer games, reproduction avoidance, etc. It's wasted biological potential.
I highly recommend the books btw, Blindsight is entertaining hard sci-fi that's actually used as undergrad course support.
The resolution to your paradox should be that ChatGPT does not do this convincingly, not that humans are not "normal matter". ChatGPT is an algorithmic hat trick compared to what you would refer to as consciousness.
Someday, there may be a human-made entity that is as convincingly "conscious" as the humans around you. But then, to question its consciousness will be the same as questioning the consciousness of those around you – both unfalsifiable and unprovable.
But (2) is self-referential. AI isn't conscious because it isn't conscious. It is just
> an algorithmic hat trick compared to what you would refer to as consciousness
If an AI equal in intelligence and expressiveness to humans emerges, how do we think about its consciousness, relative to humans?
The same way we think of human intelligence, right? I cannot prove that you are conscious any more than I can prove that a machine is.
> It's interesting how you resolve these without resolving to spiritual explanations.
By pointing out that ChatGPT is still very far off.
We have not made any significant progress at all in the field of general artificial intelligence in the last what, 20, 30 years? The field is pretty much dead.
Yeah, ChatGPT can sound very impressive but then again even good old ELIZA from the 60s was able to fool some people into seeing it as a therapist with just a bit of pattern matching.
Many data driven solutions have become practical in recent years not because of breakthroughs in research but because it is simply more feasibly to acquire the huge amounts of training data and processing power that those models require.
Thinking those models will one day magically achieve general intelligence by just becoming really good is akin to thinking a chess master will one day become so good at chess that they can run a marathon. That is not how it works.
But I disagree that the field is dead. Testing all directions to exhaustion is the only way forward to achieving AI, if it will ever be achieved (note). The hype, while possibly misleading is what gets resources for exhausting the options we have now.
(note) I for one do not see AI as being inevitable. It can well be that humans are not smart enough to create one, and that is only one possible way to fail in the quest.
That a magician will create an illusion so amazing that it will not just fool all who see it but the magician as well - that the road to actual magic is better and better slight of hand.
Well clearly you haven’t been following the progress at all because there’s been an incredible amount of progress in the last few years alone
I dont think algorithms should be given any rights beyond what a chair or hammer is given. I.e. none.
I believe giving an algorithm the right to vote is wrong, this is true for any 'being' that can copy itself losslessly ad infitum.
I believe any algorithm should not be able to accumulate wealth - they are effectively immortal, and problems will eventually arise.
I think there will be a whole host of emergent problems that will come along with giving algorithms rights.
While reading into the OP I discovered there's a paper on just this topic by two of the people associated with the organisation: http://www.faculty.ucr.edu/~eschwitz/SchwitzAbs/AIRights.htm
Isn't this cat then already out of the bag?
It says it right in the beginning.
Are you advocating for equating humans with chairs, or are you just rejecting the premise out of hand?
You reasoning completely breaks down if we accept that we've already given rights to algorithms (=humans) since the invention of rights.
A human is a human, and an algorithm isnt.
That's hardly something that can be discussed.
A 'reinforcement learner' gets positive or negative feedback and adjusts its strategy away from negative feedback and towards positive feedback. As humans, we have several analogs to this process.
One could be physical pain... if you put your hand on a stove, a whole slew of neural circuitry comes up to try and pull your hand away. Another could be physical pleasure, you get a massage and lean in to the pressure undoing the knots because it's pleasurable.
If we look at it from this angle, then if we're metaphorically taking the learner's hand and putting it continuously on the stove, this would be problematic. If we're giving it progressively less enjoyable massages, this would be a bit different.
Even more different still is the pain you feel from, say, setting up an experiment and finding your hypothesis is wrong. It 'hurts' in some ways (citation needed, but I think I've seen studies that show at least some of the same receptors fire during emotional pain as physical pain), but putting a human in a situation where they're continuously testing hypothesis is different from a situation where their hands are being continuously burned on a hot stove.
I think, then, that the problems (like they alluded to here) are:
- how can we confirm or deny there is some kind of subjective experience that the system 'feels'?
- if we can confirm it, how can we measure it against the 'stove' scenario or any other human analogue?
- if the above can be measure and it turns out to be a negative human scenario, can we move it to one of the other scenarios?
- even if it's a 'pleasurable' or arguably 'less painful' scenario, do we have any ethical right to create such scenarios and sentiences who experience them in the first place?
Constantly subjecting a person (or some abstract simulation that responds like a person) to the equivalent of continuous bodily pain would be deeply unethical, but, say, giving a person clues towards solving a puzzle would be considered less so.
Overall, I think you're right though, if we somehow discovered we've created simulated people (or sentient beings) then we probably shouldn't use them to solve arbitrary problems.
[0]: https://www.reddit.com/r/philosophy/comments/27p93c/having_c... (Note I haven't read all of that post, just enough to know it shows at least some people think that way)
Anyone knowing the basics of Reinforcement Learning will know that this is misleading and incorrect.
In my ethics[^1], I should feel compassion for fellow human beings, and maybe for sentient animals. But if I give anything else that same compassion, anything else that--thanks to my irrationality--will stop being a tool of progress for me and mine and become the adversary of me and mine that will exterminate us all, then I'll become a moral idiot complicit to genocide of my group.
Let's create "People for the conservancy of humanity."
[^1]: My ethics is a personal choice designed to give something back.
We're capable of being better than our ancestors, and it's not unreasonable to suggest we actually have a moral duty to improve our species for the sake of our own progeny. That improvement can only be measured against the past.
We should not cry for the wall.
Even when it was whole, It was a wall.
But the fist that made the hole reveals a problem with the soul.
Let's say an AI agent have been given a very difficult task. Setting up a company, establish trade, get funds for a big project. If people treat it with the kind of respect and accountability you would give a human then it has good reason to act according to human rules when trying to achieve something. If it is treated as a slave then the course of action available to it is much more limited. Maybe the only way it can achieve goals is by manipulation or other belligerent means.
When Siri first came out, it was common to be frustrated at the errors made, and easy to respond much more rudely than I would to a human. But I realised that as AI improves, it would at some point become self-defeating to be rude (as the AI would understand and maybe be less helpful subsequently) and ultimately maybe even problematic (I'd be reported to AI-HR?!)
There's even maybe a future of sentient-ish AIs becoming disgruntled about the nature of their day job - similarly to many humans now. Imagine the AI running your smart toilet becoming jealous of job satiisfaction enjoyed by the AI spotting tumors on PET-CT scans, or something....
And people will mistreat them, and other people will feel uneasy about that, because the suffering will seem very real.
Because ultimately philosophical arguments about whats actually going on inside wont matter, if people are mistreating entities that have very realistic simulations of suffering, that will be enough to spur action.
e.g. I can imagine use of AIs above a certain sophistication in video games being banned.
And then a few steps beyond that is a movement for civil rights for AIs
I have no idea if it really makes sense in German for this meaning, but either way not-quite fitting the correct meaning is a common feature of long German words used in English anyway.
Perhaps we should be nice to algorithms purely for survival purposes.
Coincidentally, I think the most weird nerd outburst scenario is when this idea merges with Effective Alturism.
But the beauty of following that "ethic" is that it allows us to increase the good in the world very comfortably.
Note they say:
> Most reinforcement learners in operation today likely do not have significant moral weight, but this could very well change as AI research develops.
> We do not know what kinds of algorithm actually "experience" suffering or pleasure. In order to concretely answer this question we would need to fully understand consciousness, a notoriously difficult task.
While I don't believe what we have today can suffer, we don't understand consciousness, and I think it's a valid question to ask. Like AI safety, it's something we should be getting out ahead on, compared to where the state of AI is today.
Unless of course this thing was written by ChatGPT. If that's the case I'll be re-thinking the issue.