All of the best AI-made software projects are also driven by experienced human software developers steering and priming the models. Does that mean the projects "aren't made by AI"?
No, it just means AI is not quite good enough yet to fully replace humans, and, so, unsurprisingly, the best results will be obtained from people who are already great at a field and who take the time to squeeze as much force multiplication out of LLMs as possible. The AI is still doing well over 95% of the significant work.
The question is is a computer with a human stronger than a computer without a human. At what point does the hybrid go from being stronger, to the human getting in the way, or steering the computer in more wrong directions that right ones, or the human not being able to keep up. Does the human add enough extra randomness to be of value for a while, even as a minor co-processor.
But I don't think it's randomness, because that would be easy to add. It's more like a different perspective on the training data, a different set of perception categories, and a different set of skills used to work with all of the above.
Those skills aren't very efficient, but they're the best we can do. We're used to their strengths but we don't like to think about their limitations.
It's completely plausible that AI will replace some of them, and not implausible it could replace and improve on all of them.
Terrence Tao's conversation with ChatGPT is very illuminating.
HN Discussion: https://news.ycombinator.com/item?id=49010345
Software devs steer and prime compilers too. Those tools don't "make the project" and nor does your so-called AI.
I prefer to judge by results.
> They themselves will freely say the AI is doing nearly everything
Of course they will.
> and that they often are not even reviewing the code
Indeed the best way to be satisfied by a chatbot's output is never to read it.
False, navier stokes was solved in one shot without steering
I think one of the significant recent proofs was basically just one person saying "solve this" "keep trying" "try harder" until it was done. I believe most of the rest can be assumed to be the human mathematicians at the very least providing some useful guidance.
I agree that it won't be that long before the AIs can do basically everything autonomously, though. I am essentially a Singularitarian when it comes to my outlook, even if not for September 2026.
The bigger question is to what extent did expert mathematicians metaprompt the model with fruitful solution strategies through their sessions finding their way into training data. Answering that question definitively is kind of important for understanding the models contribution/capability. But I feel like people want to turn this into a debate about priority and credit which is sort of secondary
E.g. in the first famous computer assisted proof (of the four color theorem) the computer only executed the resulting calculations defined from the new logic, it did not have part in the work needed to show those calculations could answer the problem nor did it come up with the actual calculations to do.
where did you "hear" this? OpenAI said they only prompted it and it solved the problem in one shot without any help