HN
Hacker News
Top
New
Best
Ask
Show
Jobs
Comment by mfa1999 | Hacker News Reader
Full thread
mfa1999
·
How does this compare to llama.cpp in terms of performance?
View on HN
solarkraft
·
MLX is a bit faster (low double digit percentage), but uses a bit more RAM. Worthwhile tradeoff for many.
ysleepy
·
On my M4 Pro MLX has almost 2x tok/s
Reply on news.ycombinator.com