On the contrary, pi + glm + DeepSeek… bliss.
Fable was a different kind of beast though. Rip.
On the contrary, pi + glm + DeepSeek… bliss.
Fable was a different kind of beast though. Rip.
I had only a few places where I did spot a difference but that difference was significant and I can imagine where people would be amazed.
What kind of "experimental stuff and complex systems" did you try that it excelled at?
On a large C codebase, Claude hallucinates constantly, and GPT 5.5 gets there are with a lot of help, but still gets things wrong.
I'm working in a 600k+ LoC codebase that has complex domain-specific logic and lots of moving parts. I find that Codex 5.5 is pretty good at surgical fixes, but does not go out of its way to explore and figure out what those surgical fixes might break. So I only use it to work on parts of the system that are pretty isolated from everything else so that risk of regression is small.
For most important work (complex, cross-domain inquiries etc.), I still rely on Codex GPT 5.5 though.