So, while an LLM might be capable of developing consciousness at a big enough scale, a human is not just an LLM, so human-like consciousness would need a different and more advanced architecture, not just a bigger training corpus.
So, while an LLM might be capable of developing consciousness at a big enough scale, a human is not just an LLM, so human-like consciousness would need a different and more advanced architecture, not just a bigger training corpus.
An LLM's temporal resolution is certainly worse, on the order of minutes to months, but not fundamentally different. By fundamentally different I mean like comparing color and time.
An LLM would have no problem living as a sloth.
So? We can still tell 1 munute from 1 second.
>An LLM's temporal resolution is certainly worse, on the order of minutes to months, but not fundamentally different.
At the moment it's not just slower, it's zero, cause they don't keep state of what they do as part of their training iirc.
LLMs have a context window, as it would be impossible to carry on a conversation with them without it. They can answer questions about when some event in that conversation happened relative to other events in it. GPT4's 32K tokens isn't human capacity, but it's not zero.
And you can already have a sense of time passing with the existing LLMs if you just feed them inputs like "X seconds passed" etc. Connect that to an actual clock, and they have a more accurate time than humans do.
I should also note that when this is tried with models with less RLHF (i.e. not ChatGPT), they get "depressed" very quickly if the only input is time passing and nothing else. I actually had LLaMA threaten me repeatedly across several such experiments.
Now the reason it's still a "maybe" then is because we would need to reasonably prove it's not a stochastic parrot.
I don't see why it matters if the "time signal", whatever it is - and you surely need one for an internal clock either way - is text or something else. The models that we have only have text inputs, so naturally it would be a token (but it could easily be a specialized non-text token like BOS/EOS if we trained the model that way). And the model can abstain from generating anything given any input - this is actually not uncommon for smaller models. GPT-3.5 and GPT-4 never seem to do it, but then again it's specifically fine-tuned for chat, i.e. always producing an output.
Long-term memory is a general problem with these things, but its short-term memory is its context window, so why would it have problem correlating events there? And for long-term memory, if it is implemented as an API under the hood that the model uses to store and query data, it would be trivial for it to timestamp everything according to the clock, no?
We're a long way from anything like an AI that thinks many times faster than humans. But I think giving AIs something like a game (or games) they can play while they're not otherwise being interacted with and just aware of the time would be genuinely useful to go some way to solving the "psychosis/depression" problem they have when their only sensory input is just time ticking over. Not necessarily the most computationally efficient, but maybe we'll find some shortcuts.
It would be a necessary, but not necesarrily (no pun intended) sufficient prerequisite.