There is quite a few on Google Image search.
On the other hand they still seem to struggle!
We now need to start using walrusses riding rickshaws
Of course by now it'll be in-distribution. Time for a new benchmark...
E.g., the pelicans all look pretty cruddy including this one, but the fact that they are being delivered in .SVG is a bigger deal than the quality of the artwork itself, IMHO. This isn't a diffusion model, it's an autoregressive transformer imitating one. The wonder isn't that it's done badly, it's that it's happening at all.
The point is never the pelican. The point is that if a thing has information about pelicans, and has information about bicycles, then why can't it combine those ideas? Is it because it's not intelligent?
https://road.cc/content/blog/90885-science-cycology-can-you-...
ChatGPT seems to perform better than most, but with notable missing elements (where's the chain or the handlebars?). I'm not sure if those are due to a lack of understanding, or artistic liberties taken by the model?
And in ChatGPT Pro.