As long as we’re seeing things like that, it’s saying a lot about the AI companies’ trust in the capabilities and reliability of their models and harnesses.
As long as we’re seeing things like that, it’s saying a lot about the AI companies’ trust in the capabilities and reliability of their models and harnesses.
I don't. I want the tools we've got now, but progressively more effective and more useful.
"You have a very strong track record of speaking out of both sides of your mouth in different settings."
If I have a consistent record of that it should be very easy for you to come up with the some examples.
There are many many more.
If you're looking for coverage that talks about what doesn't work as well as things that do then I've been persistently providing that for four years now.
Plus prompt injection, AI misuse, AI ethics... I consider all of those part of my "beat" in covering this industry.
My post about Claude's latest system prompt (and how it was likely inspired by lawsuits filed against Anthropic) was on the homepage here just a few hours ago. Is that uncritical? https://simonwillison.net/2026/Sep/2/claudes-new-system-prom...
Since we’re pulling random posts, was your blog post celebrating going out with your family while Claude wrote a large project for you an instance of your critical stance on AI capability? Have your many enthusiastic posts about increased AI capability shown a critical stance?
I think that's a good illustration of my critical stance. It ends with several open questions:
Even if this is legal, is it ethical to build a library in this way?
Does this format of development hurt the open source ecosystem?
Can I even assert copyright over this, given how much of the work was produced by the LLM?
Is it responsible to publish software libraries built in this way?
Which I later answered in another post: https://simonwillison.net/2026/Jan/11/answers/
https://dictionary.cambridge.org/dictionary/english/criticis... presents two different definitions (among several) of "criticism":
"the act of saying that something or someone is bad or a comment that says what is bad about it"
And
"writing or speech that expresses opinions or judgments about the good or bad qualities of something or someone"
I'm talking about the second here, not the first.
Give me a moment to consider your previous post.
The question you were most interested in engaging with was Does this format of development hurt the open source ecosystem, which has some overlap with the previous one in all fairness. Your answer to this is framed around the productive output of the open source community, not around the well-being of those involved. This misrepresents the question as stated. You also begin by suggesting all detractors are of lesser status and so their opinions can be ignored.
No. I don’t see it.
Yeah, this is one of the harms of Ai That I don't think is discussed nearly enough: the psychological toll it has on people who see it as devaluing their work and potentially their entire careers.
(I've been calling that "Deep Blue".)
I've been trying to find ways to help there by consistently emphasizing that these tools amplify existing expertise and, if anything, make our skills more valuable than they were before. I do believe that to be true.
My overall approach to all of this has been that I don't think it will be uninvented or banned, so the best we can do is figure out how to apply it as positively as possible and help bring by as many people along with us.
See also "slop" - I helped amplify that term precisely because it's such a great way of classifying negative applications of AI, and hopefully helping stigmatize them.
Hopefully all of the above demonstrates that, while I may not live up to your high standards, I am at least in a different category from the breathless AI boosters that infest Twitter and LinkedIn.
I hope that people who follow my work develop a deeper and more nuanced understanding of LLMs as a result.
I don’t understand why you would believe this when the opposite view is the natural one based on the history of capability increase over the last several years. The tools can now operate at an expert level in many domains and there’s no signs of the capability increase slowing. With this investment in infrastructure, surely the models today will be terrible compared to the ones 2 years from now. Moreover, the goal of these companies is to displace work, not augment it. And I have yet to see them fundamentally err in a single one of their predictions, to my great distress. You simply do not achieve a 25-30 trillion dollar total addressable market by merely augmenting workers.
(There are several reasons I disagree with this point but I digress.)
> I am at least in a different category from the breathless AI boosters that infest Twitter and LinkedIn.
You are the one whose work Anthropic plants Easter eggs of in their product releases. I don’t know what they say on X. I doubt they have the same influence. I have seen many people express incredible contempt for humanity in support of this AI revolution though.
> My overall approach to all of this has been that I don't think it will be uninvented or banned, so the best we can do is figure out how to apply it as positively as possible and help bring by as many people along with us.
We agree that it’s not going to be uninvented. I really am afraid that “us” is just a tiny minority of people who will benefit enormously from this to the detriment of most of the others.
For what it's worth, I kept engaging because you clearly had a perspective based on real issues and deep consideration.
The important point is that the system prompt here doesn’t describe the actual goal of the instructions, which (presumably) is to prevent copyright infringement [1]. This means, in turn, that the AI isn’t trusted to accomplish goals that it is instructed with. That in itself constitutes a pretty serious caveat for what we would like to use AI for.
[1] Even assuming that the goal is not to prevent copyright infringement, but instead to prevent mere accusation of copyright infringement, that’s also a directive that the AI could be instructed with. But that isn’t what they chose to put into the system prompt.
On the contrary, compliance with the system prompt would prevent that judgement.
"Claude does not reproduce song lyrics, poems, or passages from books and articles, in whole or in part — including the last lines, a chorus or hook, a melody written out note by note, or lines the person pastes in one at a time and describes as their own song."
But in fact my tests show that's not happening on works out of copyright.
Presumably if negative guidance is in the system prompt, there's a good chance that the model would happily comply if it wasn't there.