Let's judge the speed of the model when its weights are released and every inference provider on the planet offers it, so demand can spread out a bit.
It's the same topic with token budget comparisons and subscription pricing - don't people understand that this doesn't really matter for open weights models? The pricing is going to be determined by the inference providers, and until they had a chance to evaluate the model on their infra and set token prices accordingly, one doesn't really have anything tangible to compare with other open models nor with closed ones.
On the other hand I expect K3 future refinements to be massive and more efficient.
Also this arm of the discussion was about speed, not price.
Yoh have posts in this thread suggesting that Fable is 5x more expensive than Kimi K3.