If I post a question to the internal payment team's forum about a critical processing issue and some "payments bot" replies to me, should I be at fault for trusting the answer?
So to a degree, corporate politics can sort of discourage it.
That is politics. Not engineering.
Assigning a human to "check the output every time" and blaming them for the faults in the output is just assigning a scapegoat.
If you have to check the AI output every single time, the AI is pointless. You can just check immediately.
1. Check frequency (between every single time and spot checks).
2. Check thoroughness (between antagonistic in-depth vs high level).
I'd agree that, if you're towards the end of both dimensions, the system is not generating any value.
A lot of folks are taking calculated (or I guess in some cases, reckless) risks right now, by moving one or both of those dimensions. I'd argue that in many situations, the risk is small and worth it. In many others, not so much.
We'll see how it goes, I suppose.
There is a point to using LLMs. They can save time by doing a first pass. But when they do the last pass, disasters will follow.
"Debugging is twice as hard as writing the code in the first place. Therefore, if you write the code as cleverly as possible, you are, by definition, not smart enough to debug it." ~ Brian Kernighan
"We cannot solve our problems with the same thinking we used when we created them."
I'm pretty happy with the team I've built. They make solid decisions that I can trust every time. I can't say the same for the LLM.
https://www.psypost.org/scholars-ai-isnt-hallucinating-its-b...
So I asked AI to give it a good name, and it said “statistical wandering” or “logical improv”.