Show HN: Qwen3.6-35B-A3B on a 16 GB M1 Pro with SSD-streamed MoE
github.com
github.com
I've been working on a mesh environment that relies on an explicit prefix hash in the request and enables constructing a new session with a cached pre-filled system prompt specific KV-cache beyond what OpenAI-compatible APIs offer. Can you see a feature like that being supported?