Or do you have to copy paste into LM studio?
Or do you have to copy paste into LM studio?
https://opencode.ai and https://github.com/QwenLM/qwen-code both allow you to configure any API as the LLM provider.
That said, running agentic workloads on local LLMs will be a short and losing battle against context size if you don't have hardware specifically bought for this purpose. You can get it running and it will work for several autonomous actions but not nearly as long as a hosted frontier model will work.
It seems to work better for me when I tested it and Cline's supposedly adding it to the Ollama integration. I suspect that type of alternate local configuration will proliferate into the adjacent projects like Roo, Kilo, Continue, etc.
Apple adding hardware to speed it up will be even better, the next time I buy a new computer.