I use the following models:
- Qwen3.5 9B
- Qwen3.6 35B A3B
- Qwen3.8 27B
- Gemma 4 26B A4B
- Gemma 31B
- Muse Glimmer 30B
I have 48G.The MacMini is used solely for inference. llama.cpp + llama-swap.
People are saying local models are crap and serves no purposes.
I use them to help me spell check, write emails, write text messages, write JIRA tickets, write PR comments, etc.
I also use them in coding agents to complete different tasks.
I find them quite useful!