It's just not there yet. I tend to be kind of bearish on LLMs in general, I think there's a lot more hype than is warranted, and people are overlooking some pretty significant downsides like prompt-injection that are going to end up making them a lot harder to use in ubiquitous contexts in practice, but... I mean, the big LLMs (even GPT-3.5) are definitely still in a class above LLaMA. I understand why they're hyped.
I look at GPT and think, "I'm not sure this is worth the trouble of using." But I look at LLaMA and I'm not sure how/where to use it at all. It's a whole different level of output.
But that doesn't mean I'm not rooting for the "hobbyists" to succeed. And it doesn't mean LLaMA can't succeed, it doesn't necessarily need to be better than GPT-4, it just needs to be good enough at a lot of the stuff GPT-4 does to be usable, and to have the accessibility and access outweigh everything else. It's just not there yet.