They have in-house models, and the data to train even more powerful ones. The cursor team is a proper AI lab.
They have in-house models, and the data to train even more powerful ones. The cursor team is a proper AI lab.
On the contrary, it's over selling it: it's a not even a stand-alone IDE (like Zed, for instance) it's a mere fork of VSCode.
But yes Blink definitely started as a Webkit fork, and everyone would have found that laughable if someone bought that a proprietary fork of Webkit for $60B.
85% of the compute for the final model is from them, and not the base Kimi model.
Does it perform meaningfully better than the Kimi model given all that extra compute? And proportionally to the amount spent?
However it definitly isn't _just_ Kimi. The weight will be different after that 85% of extra training on top of the base model.
If those different weights are better are worse doesn't change that it's in most meaningful ways not the same as the base one.
I would encourage you to lookup their blog posts about their post training process if you want a bit more faith that they aren't running an extra 85% of compute and burning money with no-ops.
I don't think it's all no-ops. Still don't think it's a particularly relevant model/company/product.
I'll defer the reading until I see signal that they have something worthwhile. I've watched a couple interviews and used the product, neither of which impressed me.
(Only half-joking…)
I'm not super concerned about the spend to train the model, especially given that Kimi was famously incredibly cheaply made, and given what they are competing with. I don't think that's a meaningful concern.
Reciprocally, and in far more important relevant in my humble opinion: in terms of cost to run models: Composer 2.5 is easily one of the cheapest models out there. It's fantastically cheap. It's token efficiency is through the roof astronomical. I think this training for a coding specific model has yielded something incredibly special here, and I hope SpaceXLAIC isn't the only company doing this.
Composer post training is clearly very good, only second to Anthropic and OpenAI.
It does irk me a bit that they try to hide the fact that it's based on a chinese pretrained model though.
listen and learn :)
Good luck to the alt-economy of SpaceTesla though, may all our 401ks survive.
> See here https://cursor.com/blog/composer-2-5
> 85% of the compute for the final model is from them, and not the base Kimi model.
Of course they could be lying, but it seems feasible that they are adding a lot on top of this