ParentFull threaddanbst·in llama.cpp inference runs on CPU, using AVX-2 optimizations. You don't need GPU at allIt runs on my 2015 ThinkPad!View on HN