Which makes LLMs far more dangerous than idiot humans in most cases.
And… I am really not sure punishment is the answer to fallibility, outside of almost kinky Catholicism.
The reality is these things are very good, but imperfect, much like people.
And when was the last time a support chatbot let you actually complain or bypass to a human?
Let that sink in.
An LLM doesn’t make decisions. It generates text that plausibly looks like it made a decision, when prompted with the right text.
What the “LLMs don’t reason like we humans” crowd is missing is that we humans actually don’t reason as much as we would like to believe[0].
It’s not that LLMs are perfect or rational or flawless… it’s that their gaps in these areas aren’t atypical for humans. Saying “but they don’t truly understand things like we do” betrays a lack of understanding of humans, not LLMs.
0. https://home.csulb.edu/~cwallis/382/readings/482/nisbett%20s...
I don't think there's much of a difference in practise though.
I'm afraid that's not the case. Literally yesterday I was speaking with an old friend who was telling us how one of his coworkers had presented a document with mistakes and serious miscalculations as part of some project. When my friend pointed out the mistakes, which were intuitively obvious just by critically understanding the numbers, the guy kept insisting "no, it's correct, I did it with ChatGPT". It took my friend doing the calculations explicitly and showing that they made no sense to convince the guy that it was wrong.
Certain gullible people, who tends to listen to certain charlatans.
Rational, intelligent people wouldn't consider replacing a skilled human worker with a LLM that on a good day can compete with a 3-year old.
You may see the current age as litmus for critical thinking.
Humans are also very confidently wrong a considerable portion of the time. Particularly about anything outside their direct expertise
LLMs fail in entirely novel ways you can't even fathom upfront.
Id say those are the goals we should be working for. That's the failure we want to look at. We are humans.
Trust me, so do humans. Source: have worked with humans.