Cursor is already kind of useless for third party models unless you're willing to spend thousands of dollars. It's only worth it if you're going to use mostly grok/composer.
Cursor is already kind of useless for third party models unless you're willing to spend thousands of dollars. It's only worth it if you're going to use mostly grok/composer.
I basically never use the editor but the fact that there's a review UI for all the agent work is incredibly helpful.
For me, that makes the models much more usable. I've also been using GPT models a lot, as they're cheaper and less vomit inducing than Claudes text, so this is definitely bad news for me.
In particular Haiku 4.5 is rubbish, Anthropic don’t have anything in the cheap/fast part of the market.
For people fortunate enough to still be on a subscription instead of per-token billing this is less relevant, but their time will come.
The cursor model is quite nice because it allows you to switch between cheap and expensive models for different tasks
Now, I’m no expert, because I was late to the game and have only ever used Pi. But I guess Cursor is some product that’s tied to your IDE? If so, then yeah. That’s just too restrictive. I still love my IDE, but I don’t want it to be my harness too.
Claude Code works with models from other providers too. Anthropic supports this. You can configure some Claude Code environment variables to switch: eg changing ANTHROPIC_DEFAULT_HAIKU_MODEL to point to GLM Flash or Luna, setting ANTHROPIC_BASE_URL to point to api.z.ai, and making ANTHROPIC_AUTH_TOKEN the API key for your alternative provider instead.
Some instructions here:
https://docs.z.ai/scenario-example/develop-tools/claude
That said, I've not actually tried this myself, opting to build my own harness instead. And you don't know what information Claude Code might be sending back to Anthropic about how you use competing models and which models you use. I don't know for sure that they do this, but after hearing about how they used steganography in the date of harness system prompts to identify the user's location, I don't entirely trust Claude Code anymore.
Why do you say that? I've used Claude Code with DeepSeek via OpenRouter just fine.
Additionally helpful where 2 apps/repos might require running at the same time, e.g. headless web apps, or, for plugin development where the plugin might be a dedicated repo but needs to run in another app to observe changes and make it re-test itself.
I also switched from terminal to the app not that long ago, I don't find it buggy, it has access to a browser which is really helpful.
However, what I see cooking at Cursor is way more promising than the others. Things like full system understanding, their own forge, multi-repo support. I'd put them as way more visionary in the Gartner magic Quadrant ;)
For the forges, what the sibling said.
I think it's possible to have a better TUI, with something that isn't modeled after a shell prompt. Apparently, recent updates move in that direction. The GUI version of Claude Code seems to be coming to Linux, too.
Inference is cheap; Dax once said on a podcast that Opencode has a close to 90% profit margin on the openweight models it provides inference for (at 10x cheaper pricing). Those models are close in size and spec to models from large labs.
If either one of them stops doing that, customers will stop using them, because other models will become better.
In other words, there is no point in talking about inference cost in isolation.
When they outsource training to China they will be so profitable!
Honestly, _this_ was their moat more than Composer to me.
I dont know how cursor works. Everytime ive tried to use, its super buggy and has memory leaks that make my very quite PC sound like a jet engine.
It is for 3rd party models (which you get double the amount of your sub price as usage).
Of course they wont charge you extra for BYOK. how would they???
> On Teams and Enterprise plans, third-party model requests include a Cursor Token Rate of $0.25 per million tokens. This rate applies on top of model API pricing for included usage, on-demand usage, and BYOK usage.
See the link I posted above or also this one: https://cursor.com/help/models-and-usage/token-rate
Why? Written by does not mean designed by etc. There's a lot more to it.
Shit, for most things, ive got a dev agent that reads change requests from a GDoc and communicates through email with me. I can develop in my mobile
Big reason we built https://boxes.dev around the model harnesses (Codex + Claude Code), so you can bring your own subscriptions.
How long have you been in tech, outta curiosity?
Edit: great answer.
What do you mean by that? If you're being sassy about a downvote then I'll remind you that when you reply to a comment, that person can't downvote your reply. Also this is an account with barely any comments or points, I'm pretty sure they can't downvote at all.
Just a stub of a conversation.
I don’t care about the points. Just the chat, which seems finished.
But it didn't come from the person you were talking to, so editing on a sarcastic "great answer" doesn't really work.
The person you were talking to just hadn't responded within the first hour. They still might.
On the other hand, subscriptions create lock-in in a way that API pricing doesn't.
I think a more likely end is that subscription value decreases over time because API pricing gets more reasonable, but subscriptions stay because they are a good way of getting money out of people consistently.
This is why "subsidized tokens" is possibly a misnomer. Money at lower variance is worth more than the same money at higher variance. Not "subsidy" so much as reducing risk and passing some of that to a consumer.