I uploaded quantized and full precision GGUFs for local llama.cpp inference to https://huggingface.co/unsloth/Qwen3-Coder-30B-A3B-Instruct-.... Docs to run them: https://docs.unsloth.ai/basics/qwen3-coder-how-to-run-locall...
Also fixed tool calling for the 480B Coder https://docs.unsloth.ai/basics/qwen3-coder-how-to-run-locall... and made 1 million context ones as well.