My point is that it seems pretty clear that the future is in the space that OpenAI is right now. And that isn’t a bet that Apple was investing in very heavily.
That's also probably why most AI research published by Apple is about on-device inference. It's expensive to run inference servers at scale. Apple is a hardware company, so it makes sense they want to focus on what you can do on a local device (or more accurately, how they can sell you a new piece of hardware).
It's just that nobody has built one yet, I'm surprised because it's a very suitable application. But I think the cost is much higher than the current scripted models, which means there must be a payment model attached. Right now all the major voice assistants are free and I have a feeling they're all waiting to see who makes the first paid LLM-based product, and how the market reacts.
The former is really straightforward to implement with an LLM — it’s basically what an LLM is.
The latter is a whole different story.
You just prompt it with "If I'm asking you to call someone, please output only "<CALL>" and the name of the person". Then capture that keyword.
It works fine like that.
The open problem here is making it on-device, and privacy preserving. Though I’m optimistic about this, as Apple has bought up a huge number of AI startups in the last couple of years, so they are probably onto something.