Presumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part.
1,615 karma · joined February 5, 2023
Presumably the fact that they're heavily trained to reply in this way? I don't know about the rest of the paper, but this part sticks out as a really odd claim unless I'm entirely misunderstanding this part.
The language doesn't appear to be designed around supporting existing tools (either by exporting to Why3 or manually interfacing with existing tools). I'm not against shipping the whole proof or storing it on a cache server, I'm against the idea of having a LLM write it all. Even having the LLM only write proofs for subprograms that take a long time for ATPs to prove would work.
For the example in the article, the LLM had to write it exactly once without any iteration and it proved in a second, which I assume was mostly startup time. Having a LLM write 442 lines instead, which I assume also needed some iteration, is a tough sell in comparison.
Most of what I have to prove is floating-point code where a manual proof is too much of a headache to ever attempt though.
Sorry. See the edit at the top if you haven't already. I didn't realise how much it came off as a critique of you rather than a particular approach to software engineering.
It's easy to write something and have a model of what you're writing in your head that is massively different from how someone else will read it without realising, not that that excuses it.
---
I disagree with LLMs manually writing proofs without other tools doing all the work they possibly can ever being a good solution for a couple of reasons:
1. Tokens are really expensive when we have a LLM spending hours hacking aware at a proof, not to mention generating those tokens is slow.
2. The context window becomes flooded with proof work rather than work on the original problem, which will lead to a worse solution. LLMs are demonstrably worse at writing code when you continue a session on a new task instead of starting a new one.
> It is my vision that a good proof language should be fully explicit, because this reduces proof-checking time significantly.
We can cache the results and help the checker along with assertions rather than throwing out all the smart parts of the checker.
This is just about vibe-coded programs in general when the approach assumed by the article is taken. For all I know they did make an informed decision regarding the tradeoffs (which I would consider to be a poor decision).
> I just want to be able to write C#, JavaScript or whatever, and then tack on preconditions, checks and so on with the same syntax.
That's more or less what SPARK (and others) do, although specifications for large programs can become nasty.
My critiques of the language itself are not the main point, although I do still think that it's a very bad design to have a LLM waste tokens on a proof that could be written by CVC etc..
I'm not familiar with the author, I just saw the language posted the other day. I'll add a note to the top.
> These are different approaches with different trade offs.
Why would we want the tradeoff where the LLM has to write significantly more code and where the specification needs to be more complicated? If the author is aware of the state of the art then I think they made a poor choice, but that's not the point.
> Your post isn't clear, you don't go into any of these details
Bend just serves as a useful example, my general point is about how people will vibe-code a solution without an understanding of the field, leading to worse results than if they spent a little while understanding the field and then vibe-coded their thing.
The reason you don't want Blender files at all in any sort of engineering is because it's a mesh. You can't pull any useful information, even something simple like a radius, from a mesh.
This entire idea of charging or discharging your body for some physiological effect is fundamentally flawed. It is indeed true that touching the ground will change your electrical potential to that of the ground, however that's as far as it goes, the rest is just nonsensical.
Let's start with the very first part of that, specifically changing your electrical potential to match the ground. The problem with this is that the ground is 0V because we define it as 0V, not because of any sort of fundamental property. Between two points on the Earth you can see hundreds of kilovolts. So the obvious question is why should the electrical potential of a particular piece of soil where your grounding rod is installed be physiologically preferable to whatever potential your body happened to be at before?
Now lets go on to the next part of what these scam artists usually claim. That is that the Earth's electrons enter your body and do good things of some description. When you touch the ground some small charge does enter or leave your body, however that stops nearly instantly when the potentials equalise[1]. How is the Earth meant to to keep putting electrons into your body if there's no voltage? That's just simply not how electricity works.
As an aside, note that I said enter or leave. A positive static charge and a negative static charge can both easily build up in your body. Therefore any claims that specify that electrons enter your body (or leave you body) is immediately nonsensical because it could go either way.
I'll leave it to someone else to explain all the nonsensical biological claims that follow on from the already fundamentally wrong electrical side since that's not my area. I will however note that none of these sponsored papers actually show and measurable biological effect, they all just hand-wave about redox reactions.
[1]: Yes, you're coupled to mains wiring and other changing fields that are around, however that's getting too far into the weeds for a short comment. If the claim made by the snake oil salesmen was that a current going through your body is good then they'd also suggest wrapping yourself in an electrical cord to maximise the effect, or just taping batteries to your skin.
Though for the record, there was no community spread early in COVID where I am largely because there were approximately zero COVID deniers here: https://en.wikipedia.org/wiki/COVID-19_pandemic_in_Western_A...
The value would be in images reposted to social media where the website an show a badge that says it came from a certain source.