The proof is in the pudding, show me the code!
In my experience LLM are smart but sometimes inconsistent and over a long chat it might say things that are logically self contradictions… when you tell it that it confirms it.
It just seems like it lacks a consistent world view.
I don’t trust them yet. Maybe with even more scale they become better.
They act a little bit like young children, with a lot of domain knowledge.