OpenAI and Anthropic's moat is filling with cement faster than they can dig.
OpenAI and Anthropic's moat is filling with cement faster than they can dig.
If you could go out on the street of anytown and find one person using an open model, I'd eat my GPU.
Even so, I can't really run at hundreds of tokens per second which is practically table stakes for my work. Even if I did manage to run that fast, the model would probably be completely braindead and stomp all over the task.
Wish I could afford an M5 Max but I've been between jobs for months without even a single interview. Sucks to be a developer these days.
I have had very good results and compared to others it just costs pennies.
I use something similar to this https://github.com/ScotterMonk/AgentAutoFlow setup and switch between deepseek v4 to flash depending on task.