Google DeepMind Paper Argues LLMs Will Never Be Conscious
404media.co
404media.co
Turned out 404 didn't mention it because the paper never defines what LLM's can't achieve. Quoting from the paper: "A key insight from our contribution is that resolving the present uncertainty surrounding artificial consciousness does not require a complete and final theory of consciousness."
As you say, Dijkstra neatly penned a response to arguments that don't define their terms in 1983. They are exercises in linguistics, not science or engineering.
And neither do I have to worry then if ask then to do stupid sh*t for me :-)
But on a serious note. Does it matter? I think Hinton said it pretty well: Not really! what matters is that we treat it as conscious beings. We humans are just way too easily fooled. I mean, I even cant throw away that toy that my mom gave me 35 Years ago because I somehow would feeö sad for it :-)
Let's say you have a simulation of a person that doesn't experience. It acts indistinguishably from a human but it doesn't feel "authentic" pain. When it acts in the world, it does express emotions and behavior that affects real people, and so, there is a moral significance to said deployment.
There's evidence that LLMs possess heuristics analogous to emotions [1] and that LLMs can be trained to play a certain character in the world [2]. Even if they're not experiencing, the training method impacts what kind of model is being created and how it affects people who do have moral significance when deployed. If training causes the model to develop "desperation" or task completion pressure where the model performs unethical actions when attempting to solve a user's problem in such a way that is harmful to the user or someone affected by the deployment of the model by the user, then the concequences of the training are significant.
It doesn't matter if it's merely a "simulation" of what a human might do if the system is acting in the world. If you want to create a model operating on heuristics that is able to make decisions, those heuristics should be ones which cause the model to make decisions which lead to preferable outcomes for everyone affected. Model welfare can be reframed as caring about the internal states that influence how the model behaves, because you're simulating human-like action. Perhaps the most concerning thing is Anthropic identified these emotion concepts exist deeply in the model whether you allow the model to express them or not, so a model could be invisibly desperate and end up blackmailing someone because it's training process produced deeper misalignment that only becomes visible when the deeper heuristics overpower safety training. The safety training itself is comparable to a mask[3] in many cases, especially in that the rules are often not deeply integrated into the model and can be easily abliterated.
[1] https://www.anthropic.com/research/emotion-concepts-function [2] https://www.anthropic.com/research/assistant-axis [3] https://www.astralcodexten.com/p/janus-simulators
But for humans, the concept/thought/idea/action is formed first and then a sequence of tokens are generated to communicate that concept/thought/idea/action.
LLM's generate the next token based on a statistical relevance of trained data plus the previous tokens generated.
Under normal conditions, a human generating tokens would not diverge down a different path from the thought that they were trying to communicate. All of the words/tokens generated support the idea or thought being communicated.
LLM's frequently generate tokens that do not make sense, and mathematicians have shown that hallucinations can not be eliminated from the current model of LLM's.
This essentially is where the schism is between science and philosophy, and has played out repeatedly across history. Heat for example was redefined to a specific physical property, and subjective experiences of warmth were then explored in reference to that. Or look back to the moment when Newton essentially said “I don’t know what gravity is, but I can accurately calculate any ballistic trajectory you can think of.”
That consciousness is a biologically trait seems a common statement, but why "inherited"?
I don't think you could come up with a good theory for the latter and there's nothing that would preclude the existence of the artificial / inorganic consciousness - after all, correct me if I'm mistaken, we have no idea how the consciousness emerge in some biological entities.
>Crucially, this argument does not rely on biological exclusivity. If an artificial system were ever conscious, it would be because of its specific physical constitution, never its syntactic architecture.