GPT-5 goes hard on real-world programming
omerba.dev
omerba.dev
When it comes to programming, I have to keep replaying with "Nope", "No", "Again", "Wrong", "It doesn't work" and a couple of times "Do better" before it finally produces something complex that actually works.
With coding at least I know when it's wrong. The real problem is when I don't.
Since LLMs rely on patterns learned from large amounts of text, the results are relative. If the training data contained more Rust repos, that could explain why it feels stronger in Rust.
The way AI companies talk about "intelligence" now is shifting. They admit LLMs can't truly reason with the current architecture, so intelligence is being framed as the ability to solve problems using patterns learned from text, not reasoning on their own. That's a big downgrade from the original idea of AI reaching human-level thinking and developing AGI.
Also, my understanding is that since Microsoft invests in Copilot, it doesn't want ChatGPT to get better at coding. Instead, it wants it to get better at being a lawyer.
> this zig code is trash
https://lobste.rs/s/1qr3zy/gpt_5_goes_hard_on_real_world_pro...