This model did suspiciously well on the pelican test. Simon has mentioned something along the lines of AI companies cheating his benchmark, and using other random absurd prompts to thwart it.
The real pelican test is about if a text producing model can spit out SVG code that renders a good vector illustration, which is a much harder (and sillier) task.