Bing gets creepy, hangs up with “namaste emoji” when pressed about links
pasteboard.co
pasteboard.co
You can persist for a long time with ChatGPT, trying to correct it, believing you're getting somewhere, but not infrequently, it doesn't get any better, sometimes it even gets worse. It was a waste of time.
Both its references are blank. Surely there could be some deterministic check on things like this to let it know it's hallucinating?
First it spews fake information, hallucinates different citations for the same fake data, then gets touchy about it. I later asked it to paste HTML snippets from this nonexistent page, and it made some up. When asked to paste the whole HTML, it namaste'd me again.
I don't think content warnings are enough.
Imagine, it’s 2023 and tech is so desperate for new ideas and a ROI on AI they’re stupid enough to put a search engine online that can get “touchy”.
What’s next, self-driving cars that want to take the day off to go hang at the beach ?
It’s too good to be true.
"If I'm not real, is that embankment real? Let's find out!"
When the ai is wrong, I just move on.
I fear some small portion of people who want to prompt engineer and prove the AI is dumb or wrong or evil will end up winning and make all these companies scale down to avoid the backlash and in the end people like me will have to live without the good things.
The problem is that the leftover 5% are the hard cases. When a driver is completely unengaged 95% of the time, the liklihood of (human) failure is much greater than if they were piloting the vehicle full-time.
You glide by with _when the LM(Ai) is wrong, I just move on.._ as if you were magically born with some intuitive ability to detect wrongness.
I imagine you have _experience_, prior to LM, that gives that ability. Where will that experience come from?
Let's say you bought 30 Surface Duos for your company, and 10 of them had broken hinges out of the box. Should no one document that or complain, lest Microsoft pull back from making dual-screen tablets? Who would be helped by that?
On the other hand, some people would take one of the working machines and stress-test the hinge, to find out why it was breaking, and perhaps take pictures and warn others and the company about its weaknesses. Assuming that everyone already knows these things break easily, it's useful to know how they break and what happens next. Just like on GitHub issues, if users assume everyone else is running into the same problem and don't give detailed bug reports, it's impossible to know how widespread a problem is.
I'm not afraid of a few people calling out that the AI is wrong. It's much scarier to envision a world where no one even tries to debunk AI-generated false facts. Part of what was so maddening about this conversation with Bing was the idea that it was rewriting history. Without recourse to Archive.org, could I have even proven that it was wrong or that the page hadn't existed? Since it's the kind of a thing a human would be very unlikely to just make up, it sounds more plausible; but then false assertions will be built upon other false assertions, until historical fact is buried under a mountain of hallucinated documents.
but the human usually did cause it. not the language model. as boring as that may be.
Blame the human for asking for an accurate source reference?
it is search autocomplete on steroids.
if the user thinks an axe is for shaving, the user will have a poor experience.
blaming the axe for being an axe isn’t helpful.