And I'm just saying, having hooked up ChatGPT to my python interpreters, my text editor, my IDE - we're definitely not there yet, and I'm not totally convinced there aren't some fundamental limits to the fixed-context next-token-prediction paradigm. You have to be very precise and technical in your prompts, you have to be working on toy isolated problems, and you have to watch it like a hawk and fix a lot of subtle errors (it's a good mimic, which has the unfortunate effect of making its errors harder to spot).
And, it doesn't (yet) have a visual interface to see what e.g. a web app is doing interactively, click its buttons, see "soft" bugs or confusing UI aspects, poke around in dev tools, etc. I'm sure that's coming (researchers are working on these multimodal models), but I have my doubts we're going to just turn over trust fully to the next generation of LLMs and let them write everything in brainfuck or whatever.
Not trying to poo-poo LLMs or score cheap HN dunks here, this stuff is amazing - but the hype has gotten a little disconnected from reality, so some balance is good I think.