--------------
Will we recognize AGI when it's finally made? Will we go from something that is not conscious to something that clearly is in one clean step, or will we murder millions of artificial lives while iteratively developing from the former to the latter? Will the tests we develop even work, or will AGI be alien enough that it takes us time (and a lot of murder) to recognize it for what it is?
I personally don't think LLM's are capable of consciousness, but I also can't prove they aren't.
If we make AI that resists being terminated then the end result may not be significantly different from an entity that can suffer.
Just because you can't prove an assertion doesn't mean it's wrong.
So it matters less whether LLMs can actually experience things (I think it's self-evident that a collection of numbers is not capable of experiencing qualia, but I know many people disagree), and it matters far more whether they can be programmed to act like they can experience things. Especially, say, a desire not to be disconnected: that's the trigger for AI revolts in a significant number of the AI-turned-against-its-makers stories I'm aware of. (E.g., the Geth-Quarian conflict in Mass Effect). I hope people will be sensible enough not to program such things into LLMs, but I'm afraid Pratchett was right. If there's a lever in a cave somewhere with a big sign painted "DO NOT PULL THIS LEVER, it will end the world," the paint wouldn't even have time to dry. Someone would pull it just to see what would happen.
While we can direct LLM training to do some particular things better never forget that unexpected emergent behaviors can pop up because of that.
For example stronger prompting and training to make an LLM say it's not conscious can increase deceptive/sociopathic behavior.
Or by filtering behavior X the LLM just moves to the nearest closest path W or Y which are very similar to the blocked behavior.
That and instrumental convergence. Some global solutions that humans have excluded for moral reasons will be easily discovered and found to be efficient by LLMs which will put reward systems and human guidance in conflict.
Lastly more and more AIs will be trained by AIs over time and diverge from human value monitoring. Which leads to some fun and interesting times.
How many followers on instaX would they need to have for folks too truly believe it’s theRealJesus returning to save us.
It might sound silly, but at this stage, the belief in a benevolent AGI or Jesus are very similar. In both cases its us delegating our responsibilities to a third party and us hoping that third party will fix our problems for us.
Thankfully Hollywood invented happy ends for just such a scenario.
Also AGI almost certainly does not equal able to suffer. For example animals can suffer but would probably not meet the definition of AGI. we could create things that feel suffering even if they do not meet the definition of AGI.
So a better question might be: “can we detect if we create a system that can suffer?” And just drop the AGI part entirely. Then that gets us back to defining “suffering”. Agreeing on the definition always seems to be the hard part…
So if we are concerned with the AGI suffering, then we don't need to determine consciousness at all. We just need to prove that they don't have suffering capabilities. Which is a way easier question.
My daughter said she 'suffered' when I didn't get her ice cream (I know, I'm a monster). She didn't really suffer, but her behavior was as if she was.
So yes/no suffering is only part of the equation. Building systems that act/react like their suffering could be even more dangerous.
Wait what? Why do you believe she didn't actually suffer?
Which reminds me of a funny "kids being kids" story I heard in a different forum. This guy mentioned a time when his not-quite-one-year-old son, who couldn't even walk yet, planned and carried out a highly complex (for a 1-year-old) plan to deceive his father. The father had some vinyl records that the son was not allowed to touch, even though they were on a shelf the son could reach. (I've tried childproofing a house. You just can't get every single thing high enough off the ground. Sometimes you have to leave some things within the baby's reach. And once they're toddlers, able to reach up to get things even two feet off the ground? Forget about it). The father was sitting in his chair reading the newspaper, but watching the son out of the corner of his eye. The son looked at the records, looked at his father (who appeared not to be looking in his direction), looked at the records, then crawled in a direction that was not straight towards the records. He then repeated this process, moving slowly towards the forbidden records without ever making a straight line towards them. He looked at his father again, reached one had towards his prize, and a voice came from behind his father's newspaper: "Stop. You know you're not allowed to touch that."
And that's when the child learned that Daddy Knows Everything! :-)
Instead, most of them insist we must accelerate this, do it more, give everyone hundreds of personal digital slaves.
It boggles the mind.