Any plans for being able to run the entire thing locally with local models?
llama.cpp - https://github.com/ggerganov/llama.cpp/blob/master/examples/...
or
mistral.rs - https://github.com/EricLBuehler/mistral.rs/blob/master/docs/...
lmstudio and ollama use llama.cpp underneath. cut the middle man