That's true but size of LLMs has been strongly correlated with their "intelligence".
And the "business" obvious is still doing that but the science and implementation has be realizing that this just isn't true. They're not getting AGI out of a single LLM by itself.
There's a lot to make efficient, but it should be clear to everyone that just throwing compute at larger models isn't going to magically make it rain.