The gap is huge and im tired of reading these articles constantly
gemma4-26B (#7)
qwen-3.6-27B (#9)
Certainly the gap is closing but I feel it still makes more sense to pay pennies to run the full sized open models hosted on much better hardware.
Overall I was very impressed with its open box reimplementation. I remain of the mind they are widely underrated.
You obviously need a 32GB card to run that (or a 64GB Mac), and realistically anything lesser than an RTX 5090 is going to be too slow for practical use.