This is great, thanks for putting this together.
Haven't followed it through yet, but does this model run successfully on an iPhone?
My 9 year old ran a Qwen 0.6B model using ollama quite well, anything else was too slow to offer a good UX.
Haven't followed it through yet, but does this model run successfully on an iPhone?
My 9 year old ran a Qwen 0.6B model using ollama quite well, anything else was too slow to offer a good UX.
I was thinking there was a fourth grader out there deploying models when at that age I was still learning multiplication tables.
[0] https://llm.mlc.ai/docs/deploy/ios.html#bring-your-own-model