The smarter 27B is so fast with MTP I've found I really don't need the 35B-A3B. You get around 70tk/s on a M5 Max lowering to around 40tk/s at higher context sizes.
I get 20 tokens/s on an M4 Max (larger GPU).
Go here: https://huggingface.co/unsloth/Qwen3.8-27B-GGUF
See the right hand side panel, you see a whole palette of quantizations and their respective sizes. Should give you an idea. Note that these are not the only quantizations available.