For those of us who don't want to pay for Gemini tokens, what would be the best local LLM to use for this?
(A model that can run reasonably well in a ~24GB MacBook.)
(A model that can run reasonably well in a ~24GB MacBook.)
So for the parent's Macbook question `gemma4-26b-mlx` should work well.
For you with 24 GB VRAM, `gemma4-26b-a4b`. I tried higher VRAM models and they slowed down while still doing just as well or slightly worse.
If someone else tests and finds a better performing model though, please update me here, I'd love to try it.