To run Code Llama locally, the 7B parameter quantized version can be downloaded and run with the open-source tool Ollama: https://github.com/jmorganca/ollama
ollama run codellama "write a python function to add two numbers"
More models coming soon (completion, python and more parameter counts)