I have the feeling that LLMs are effectively running on dream logic, and everything we've done to make them reason properly is insufficient to bring them up to human level.
I have the feeling that LLMs are effectively running on dream logic, and everything we've done to make them reason properly is insufficient to bring them up to human level.
And if you don't prove your code, do you not design at all then? Do you never draw state diagrams?
Every design is an informal proof of the solution. Rarely I write formal proofs. Most of the time I write down enough for myself to be convinced that the desing solves the problem.
As a exam grader, you can easily tell when a student has the mindset of "solving a problem" but made a mistake, and when they had the mindset of "looks like it solves the problem" and just wrote some stuff.
> Most of the time I write down enough for myself to be convinced that the desing solves the problem.
Again, why do you assume we aren't doing the same thing with LLMs?
1. Spec given
2. Ask LLM to write a bunch of design documents based off of spec
3. Ask LLM to identify edge cases
4. Ask LLM to device edge cases in to a test plan involving N tests
5. Ask LLM to write tests
6. Ask LLM to write commented code
7. Ask LLM to run tests on code, and determine on failing tests if test or code is wrong, go back to the appropriate step to fix test and/or code.
Whenever I hear someone here on HN imply that the only way to code with an AI is via vibe coding I just die a bit more inside.
It was a response to you saying: "Im not going around trying to prove my code correct when I write it manually."
How did you manage to forget what you wrote previously?
Also, in this post you are now suddenly taking the exact opposite position, contradicting your previous point.
And I have not made any statements about how you use LLMs, only about how the LLMs produce code. All statements about how you use LLMs have been made by you, not me. I haven't discussed it since it is not related to the arguments, which are: 1) whether LLMs are goal-oriented and 2) whether humans and LLMs both merely maximize plausibility when writing/generating code.
Both claims that you made. Note, however, that if you are correct in your own points, then you should indeed be able to "just dump out code without any process in between". So if anyone is claiming this, it's you.
This definitely matches my experience of talking to AI agents and chatbots. They can be extremely knowledgeable on arcane matters yet need to have obvious (to humans) assumptions pointed out to them, since they only have book smarts and not street smarts.
What they lack is multi turn long walk goal functions — which is being solved to some degree by agents.
You could build layers and layers of LLMs watching the output of each others thoughts and offering different commentary as they go, folding all the thoughts back together at the end. Currently, a group of agents acts more like a discussion than something somewhat omnipotent or omnitemporal.