Humans definitely create new information. (Well, at least some humans do.)
Humans definitely create new information. (Well, at least some humans do.)
Humans can observe new information, but that's obviously not that unique. We can reason about existing information, creating new hypotheses, but that is arguably a compression of existing information. When we act on them and observe the effects they become information, but that's not us creating the information (and LLMs can both act on the environment and have the effects fed back to their input to observe, so it's not really unique).
There is this whole field of art, but art is constantly going through a mental crisis whether anyone is creating anything new, or if it's all just derivations of what has come before. Same with dreams, which appear "novel" but might just be an artefact of our brain's process that compresses the experiences of the day.
More information: https://www.nature.com/articles/s41586-023-06600-9
you organize them into piles of ten
this created new information
LLM's are basically lossy compressors, they decrease information entropy.
Humans increase information entropy whenever we do anything.
(Some) information entropy is potentially very valuable, so in effect LLMs destroy value by design, while humans can potentially create value. From an economic point of view humans can never be replaced by an LLM.
Let's say we could somehow train an LLM on all written and spoken language from the western Roman civilization (Republic + Western Empire, up until 476 AD/CE, just so I don't muddy the experiment with near-modern timelines). Would it, without novel information from humans, ever be able to spit out a correct predecessor of modern science like atomic theory? What about steam power, would that be feasible since Romans were toying with it? How far back do we have to go on the tech tree for such an LLM be able to "discover" something novel or generate useful new information?
My thought is that the LLM would forever be "stuck" in the knowledge of the era it was trained in. Something in the complexity of human brains working together is what drives new information. We can continue training new LLMs with new information, and LLMs might be able to find new patterns in data that humans can't see and can augment our work, but the LLM's capability for novelty is stuck on a complexity treadmill, rooted in its training data.
I don't view this ability of humans as some magic consciousness, just a system so complex to us right now that we can't fully understand or re-create it. If we're stochastic parrots, we seem to be ones that are magnitudes more powerful and unpredictable than current LLMs, and maybe even constructed in a way that our current technology path can't hope to replicate.
If you took millions of genius-level immortal humans with all the same Roman data but had them sit in a blank, empty room with their hands tied and simply discuss philosophy for eternity, I'm certain that they would not be able ever spit out a correct predecessor of modern science like atomic theory. Perhaps they could spit out billions of theories including the atomic theory as well, but they would have no data to presume that the atomic theory is more relevant than any other. Extensive information processing can squeeze out every last ounce of knowledge from some data, but anything that isn't in that data can't be acquired by mere thinking about it. On the other hand, if you gave some "LLM++" the ability to toy around with reality and attempt all kinds of experiments to test various hypotheses, then I wouldn't assume that it would not be forever stuck in the knowledge of the era it was trained in.
Funny example - depending on how close the predecessor has to be, the answer is maybe: https://en.wikipedia.org/wiki/Atomism