Not all LLM models are remote. You can do just fine with local ones.
There’s so many projects that just seem to call out to the OpenAI api even though they say something silly like “99% local.”
For me, I look for project authors to describe their work as all on device pretty early on in my filter.
Not that calls to cloud AI are bad or something, but just have very different use cases.
That used to be the case for lots of business-facing products that wanted to capitalize on the hype quickly a year ago but I think the dust has settled (?)
For developer tools, however things like llamafile are pretty much the standard. Not to mention the pain of maintaining multiple keys, different response formats etc.