If it is good, we call it "creativity."
If it is bad, we call it "hallucination."
This isn't a bug (or limitation, as the authors say). It's a feature.
If it is good, we call it "creativity."
If it is bad, we call it "hallucination."
This isn't a bug (or limitation, as the authors say). It's a feature.
Just because those hallucinations sometimes randomly happens to be right, people concluded that being wrong is the exception, while being right is somehow the rule.
It's like when people read [insert millenias old text here], finds a part that happens to illustrate something in their life today and conclude that it is a prophecy that predicted the future.
The meaning/truth in those is nothing more than a cognitive bias from the mind of the reader, not an inherent quality of the text.
Sorry, somewhat trite and unfair, but, if there is a gambling-like dopamine reward cycle occurring, then the users would have a hard time being truly objective about any productivity boost in total. They may instead focus on the 'wins', without taking into account any overheads or 'losses', much as a gambler would do.
E.g. a search engine can give you zero useful results, and you can fine tune your query and still get nothing after scrolling through pages of results (Do people really take the losses into account when using search engines?) I find prompt engineering with LLMs more useful because you get nudged in interesting directions, and even if you come away with no direct results, you have more of an idea of what you are looking for. Maybe lateral thinking is overrated.
In terms of what we can expect of future improvements, I think it’s overly optimistic to expect any kind of super intelligence beyond what we see today (that is, having access to all the worlds publicly available information, or rapidly generating texts/images/videos that fall into existing creative patterns).
I suspect that more creative intelligence requires an extremely fine balance to not “go crazy”.. that is, producing output we’d consider creative rather than hallucinations.
I think getting this balance right will get exponentially harder as we create feedback loops within the AI that let its intelligence evolve.
And it’s entirely possible that humans have already optimised this creative intelligence feedback loop as much as the universe allows. Having a huge amount of knowledge can obviously benefit from more neurons/storage. But we simply don’t know if that’s true for creative intelligence yet
We’re already well past that point. Why? Because saying incredible things about AI attracts VC money.
If it isn't a bug, it dam well isn't a hallucination, or creativity.
This is a deeply integrated design defect. One that highlights what we're doing (statistically modeling lots of human language)...
Throwing more data against this path isnt going to magically make it wake up and be an AGI. And this problem is NOT going to go away.
The ML community need to back off the hype train. The first step is them not anthropomorphizing their projects.
> Imagine a piano keyboard, eighty-eight keys, only eighty-eight and yet, and yet, new tunes, melodies, harmonies are being composed upon hundreds of keyboards every day in Dorset alone. Our language, Tiger, our language, hundreds of thousands of available words, frillions of possible legitimate new ideas, so that I can say this sentence and be confident it has never been uttered before in the history of human communication: "Hold the newsreader's nose squarely, waiter, or friendly milk will countermand my trousers." One sentence, common words, but never before placed in that order. And yet, oh and yet, all of us spend our days saying the same things to each other, time after weary time, living by clichaic, learned response: "I love you", "Don't go in there", "You have no right to say that", "shut up", "I'm hungry", "that hurt", "why should I?", "it's not my fault", "help", "Marjorie is dead". You see? That surely is a thought to take out for a cream tea on a rainy Sunday afternoon.
https://abitoffryandlaurie.co.uk/sketches/language_conversat...
LLMs are just generating tokens. Hallucination perpetuates an unhelpful anthropomorphization of LLMs.
Users see it as a machine artifact.
Isn't this the difference between a human and an LLM?
A human knows it's making an educated guess and (should) say so. Or it knows when it's being creative, and can say so.
If it doesn't know which is which, then it really does bring it home that LLM's are not that much more than (very sophisticated) mechanical input-output machines.
Also, LLM-s could report more statistical measures for each answer and external tools could interpret them.
Which is still very useful for a lot of things. Just maybe not things to which value is assigned based on how efficient and correct the answer is. Like you can have GPT make a marketing campaign for you, or you can have it design all the icons you need for your application UI, but you can’t reliably make it wrote high performance back-end code without having humans judge the results. Similarly you can’t use it to teach anyone anything, not really, because unless you’re already an expert on the subject being taught, you aren’t likely to spot when it gets things wrong. I guess you can argue that a lot of teaching is flawed like that, and you wouldn’t be wrong. Like, I was taught that the pyramids was build by slave labour, even after the archeological evidence had shown this to be likely false. But our text books were a decade old because our school didn’t really renew them very often… in such a case GPT might have been a more correct teacher, but the trick is that you won’t really know. Which is made even more complicated by the fact that it might teach different things to different students. Like, I just asked ChatGPT 3.5 who build the pyramids in 3 different prompts, in one it told me it was ordinary people. In the others it told me it was mostly skilled labour under guidance of “architects” and “engineers”. Still better than teaching us it was done by slave labour like my old book, but the book was still consistent in what was considered to be the truth at the time.