OpenAI and Anthropic's moat is filling with cement faster than they can dig.
If you could go out on the street of anytown and find one person using an open model, I'd eat my GPU.
Even so, I can't really run at hundreds of tokens per second which is practically table stakes for my work. Even if I did manage to run that fast, the model would probably be completely braindead and stomp all over the task.
Wish I could afford an M5 Max but I've been between jobs for months without even a single interview. Sucks to be a developer these days.
I have had very good results and compared to others it just costs pennies.
I use something similar to this https://github.com/ScotterMonk/AgentAutoFlow setup and switch between deepseek v4 to flash depending on task.
However, Amazon was not racking debt the way these companies are. Both their behavior and financials were miles apart from these ai companies.
Compare that with how I pay $200 a month for Claude and am still hitting the limits with any sort of sustained usage. They even have a special usage limit for Sonnet to prevent you from using too much of that either.
I'm super frustrated with how slow DeepSeek is though. And it's not nearly ready to be unsupervised for long periods of time like Claude is. Just this morning I left Fable 5 unsupervised for about eight hours straight. Single turn. DeepSeek often gets even much shorter turns wrong, so I wouldn't trust it with anywhere near that length of time alone. Not to mention it'd get so much less done because of how slow it is.
Also, did you use an LLM to correct your grammar after you posted? Lol
> I'm super frustrated with how slow DeepSeek is though. And it's not nearly ready to be unsupervised for long periods of time like Claude is.
Tradeoffs ;). One thing I'm doing is to make my flows properly available on my phone, so I can run and supervise things wherever I may be.