Full threadjmorgan·That's fast. It's exciting to see more ways to run these models locally. How does this compare to llama.cpp – both in speed and approach?View on HN