https://tools.simonwillison.net/markdown-svg-renderer#url=ht...
https://tools.simonwillison.net/markdown-svg-renderer#url=ht...
I get people burning out on the pelican SVG test alongside the rest of the AI burnout, but I guess for myself I'm just choosing to keep enjoying it while I still can.
I agree to rednb that at this point it feels like rather obvious brand building, but also, I agree with you that some value is in it.
It does not feel all that authentic though, and it's good to react allergically to lack of authenticity. Bad for a lot of business models, but good for humanity.
> It does not feel all that authentic though, and it's good to react allergically to lack of authenticity. Bad for a lot of business models, but good for humanity.
I hope SimonW keeps them coming.
But why is this an indication of literally anything else?
Not only does it give you a super easy-to-grok understanding of the model quality just by looking at the image, but when you compare tokens and costs (both input and output), you really get a good, simple COST x QUALITY evaluation across models.
Simon explains it well: https://simonwillison.net/2026/Jul/16/kimi-k3/#what-can-we-l...
Simon, you should put up a summary table page that you update after every release.
Sponsored blogs and paid newsletters are after all, notoriously poor at subsisting on silence :)
Marketing, apparently, is a property of URLs rather than outcomes :)
They're easy enough to skip - click the little "-" icon and you'll collapse the entire sub-thread.
It's a decent heuristic because the better models generate better pelicans. That's all. Nobody sane is going to make a bet on a model based on a pelican. But it's cool, it's tradition by now, and it's a semblance of a good first impression for new models.
thank you.
But you bet my ass I check everytime to have a look at see how that pelican looks, it's just a fun check and also interesting to see the cost/results.
I feel like the human brain massively overweights negative feedback over positive, and that's even after accounting for the fact that internet discourse tends to mostly surface negative comments (whereas the enjoyers stay silent). By default I have to try hard not to take it personally whenever someone leaves a negative comment about my work.
So just doing my bit to say I appreciate your commentary on so much of the fast-evolving AI landscape. Helps me orient :)
Consensus has an unfortunate habit of beginning that way.
And outrage at the unimpressed is.. a curious standard for a sponsored blogger.
All of the pelicans so far have had really weird flaws / quirks so I am always a little interested to see how well these models perform at this task, since I've seen all the past pelicans and have some anchoring.
Seeing a truly flawless pelican would tell me that the model has true visual reasoning capabilities as well as good taste.
The 3.6 Flash pelican is just about the best I've seen.
The fish and the cap where always added when I asked an llm to improve it's first attempt.
This continues the trend in LLM progress of better=more stuff
Edit: I wonder if this is a function of the reasoning training, where more tokens/ stuff is rewarded.
Does this say anything about the model? I meant the underlying attention/pattern it took for Gemini 3.6 Flash to create this SVG.