There are also prebuilt binary archives for just about any distribution and inference backend for the latest github release:
https://github.com/ggml-org/llama.cpp/releases
No need to compile unless you really need to.
https://github.com/ggml-org/llama.cpp/releases
No need to compile unless you really need to.