Only it damn well isn’t. Anywhere. Not even patient reports.
The problem with AI is if it’s right 90% of the time but I have to do all the work anyway to make sure it’s not one of the 10% of times it’s extremely confidently wrong, what use is it to me?
How would you even know to evaluate that?
If an LLM gives you a response to that question, how do you know if its right or wrong without already knowing the answer or verifying it some other way? Is everyone just assuming the ai answers are right a majority of the time? Have there been large scale verification of a wide variety of questions that I'm not aware of?
I purchased a small electronic device from Japan recently. The language can be changed to English, but it’s a bit of a process.
Google’s AI just copied a Reddit comment that itself was probably AI generated. It made up instructions that are completely wrong.
I had to find an actual human written web page.
The problem is with more and more AI sloop, less humans will be motivated to write. AGI at least the first generation is going to be an extremely confident entity that refuses to be wrong.
Eventually someone is going to lose a billion dollars trusting it, and it’ll set back AI by 20 years. The biggest issue with AI is it must be right.
It’s impossible for anything to always be right since it’s impossible to know everything.
Have you tried search in ChatGPT with o4-mini or o3?
Q1.1: "do most developers do code reviews before testing"
A1.1: essentially "yes most before..."
Q1.2: "do most developers do code reviews after testing"
A1.2: answer was essentially "yes most after..."
The 2nd set of questions was related to building a retaining wall. I've never used the type of wall block with a center notch and groove, only the type with a lip on the back. The center type creates a wall straight up but lip in back leans back a little. I was curious if one was more stable than the other:
Q2.1: "is retaining wall block with center groove more stable"
A2.1: essentially "yes, block with center groove is more stable..."
Q2.2: "is retaining wall block with back lip more stable..."
A2.2: essentially "yes, block with lip on back is more stable..."
When I finally found an engineering website with real details, the answer was that the lip on back was a little more stable due to the resulting angle of the wall. It also emphasized that the lip and the center groove are not factored in to stability calcs at all, they are only for alignment, gravity and friction of the blocks is what holds it in place.
Try coding with Claude instead of Gemini. Those that do tell me it is well beyond.
Look at the recent US jobs reports--the draw down was mostly in professional services. According to the chief economist of ADP "Though layoffs continue to be rare, a hesitancy to hire and a reluctance to replace departing workers led to job losses last month."
Of course, correlation is not causation, but everyone white collar person I talk with is saying AI is making them far more productive. It's not a leap to figure out that management sees that as well.