[1]: https://pluralistic.net/2024/04/24/naming-names/#prabhakar-r...
[1]: https://pluralistic.net/2024/04/24/naming-names/#prabhakar-r...
Same with a junior dev. They don't write long form spaghetti because they're trying to write more LOC. They do it because not doing it is hard, literally above their pay grade.
I use LLM every day, but they're still completely awful at architecture. I don't think this clear lack of ability is some conspiracy.
I was only able to do that after I had solved multiple related problems in different places and started introducing subtle bugs by accident / had difficulty detecting all edge cases
I've noticed whenever I use LLMs they introduce the same kind of thing but at much smaller scales than I would. They often suggest solving the wrong problem when I prompt them to diagnose specific bugs too. Usually opting for a shortcut that introduces its own issues and ironically calling the proper direction "too complex" when it's really not.
Maybe currently not. But we will never be able to know, as models are undeterministic and benchmarks are kind of scams. When you cannot prove that something gets worse, rest assured companies will to it.
Then when competition really settles to monoploy or duopoly, like it always does in bigtech, it is really difficult to prove Anthropic and OpenAI enshittify their models. Same with Opus 4.6, that suddenly was worse when 4.7 came out.
I'm not following. Free markets and China will make sure that models stay bad!? Why didn't these entities already stop the progression we've seen? What has been motivating them all this time that has preventing them from stopping?
Keep in mind this was all science fiction just a few years ago.