[1] https://huggingface.co/TheBloke/CodeLlama-13B-Python-fp16
[1] https://huggingface.co/TheBloke/CodeLlama-13B-Python-fp16
`ollama run codellama:7b-instruct`
https://ollama.ai/blog/run-code-llama-locally
More models uploaded as we speak:
my questions were asking how to construct an indexam for postgres in c, how to write an r-tree in javascript, and how to write a binary tree in javascript.
ollama pull codellama:7b-instructI was up and running from clone/build-from-scratch/download in ~5m.
It's running on my M1.. it knows WebGL JS APIs better than I do, makes a passable attempt at VT100 ascii art, and well, should read more about Wolfram Automata, but does seem to know Game of Life!
Thank you so much!
Have had some good success with the instruct model:
codellama:7b-instruct
using <language> write me a <thing>
it's managed to spit out code, rather than "write a traversal function".
Any plans to add the 13B quant models?
This should be fixed now! To update you'll have to run:
ollama pull codellama:7b-instruct[1] https://github.com/jmorganca/ollama/blob/main/docs/api.md
curl -X POST http://localhost:11434/api/generate -d '{
"model": "codellama",
"prompt":"write a python script to add two numbers"
}'GGUF seems not optimised yet, since quantizing with a newer version of llama.cpp supporting the format fails on the same hardware. I expect that to be fixed shortly.
For inference, I understand that the hardware requirements will be identical as before.