what, exactly, were the failures?
coding assistants are not visual-first, they are code-first. so it makes sense that they excel more at coding.
coding assistants are not visual-first, they are code-first. so it makes sense that they excel more at coding.
I agree that on pure code they are OK, but I thought maybe I am missing something?
There's a reason OpenAI didn't release Codex-only variants of GPT 5.4 and up.