These benchmarks seems to suggest that Flash Next performs best on Low thinking compared to the same model on higher thinking levels?
And that Qwen 27b outperforms both when set to high?
What quants are being compared here?
And that Qwen 27b outperforms both when set to high?
What quants are being compared here?
The exact quant is mentioned on model page[0] IQ3_S, and I think 27b was Q4, via Ollama, the one that fits on a 3090 24GB
[0]: https://aibenchy.com/model/qwen-qwen3-8-flash-next-low/