Finding counterexamples is low-hanging fruit, the automation of which isn't shocking.
Finding counterexamples is low-hanging fruit, the automation of which isn't shocking.
> Finding counterexamples is low-hanging fruit, the automation of which isn't shocking.
It's not good to be confidently wrong the way you're being.
We've then improved that through systems similar to prolog intentionally searching a tree.
Then systems added heuristics for which paths in that tree are likely to be taken.
The LLMs are just using slightly more accurate heuristics for this task.
But the real measure of understanding are tasks that are not so strictly constrained.
I won't really care that it didn't have a "real measure of understanding". I'll care that it has made an even more dangerous technology, which needs work to make it aligned.