With 5 active sessions going nonstop? That seems like a pretty important qualifier.
I can pretty easily burn through my weekly quota over several agent coding hours with minimal supervision when tasked with some pretty large but well-planned refactors.
I'm also creating a free platform that replaces extremely out-of-date software, some of it only available with mutli-million dollar contracts, to help medical physics professionals with cutting-edge radiotherapy devices used to treat cancer.
https://brynnbateman.com/ for a list of projects
And out of curiosity, how do you automate testing the porting in the browser that's actually playable etc? And aren't you a bit scared of hosting and serving the "hairy bits" such as full assets? Nice job anyway!
I automate testing playable parts in browser by adding a dev mode that allows text commands for everything instead of having to rely on clicking UI or 3D elements. It can see the full game state in JSON and interact in any way via commands.
And re: IP - I just accept that I might get a C&D any day and have to take it all down. I'm careful to not accept a single penny for any reason and don't even have Patreon. Usually monetizing is what makes IP owners unhappy. And for the Pokemon MMO I just don't advertise it anywhere meaningful since they'll C&D the second they see it. Largely made it for my nephew and we play it together.
Then, once I go over, API pricing racks up FAST!
i used for work where i did less and it quickly reaches thousands if you're not careful. i can already see what some will say: skill issue et cetera - whatever.
> Not everything has to be done by the most expensive one.
Ok then.
$5/days is ~330 Mtok/day, that’s a nontrivial amount of work, and none of the gpts are more efficient than deepseek at $/task if deepseek meets your quality bar.
OpenCode currently offers 60 USD API credits at 10 USD per month (OpenCode Go) and have even doubled it temporarily as a promotion.
Effectively you can get Deepseek for 1/12th the already ridiculous cheap API price.
Per the open code zen pricing page[1], it appears that the token prices are the same, but their cache is 10x more expensive?
Deepseek 0.14/0.28/0.0028
Pro Opencode 1.74/3.48/0.145
Deepseek 0.44/0.87/0.0036
For flash the input/output is the same, but the cache difference is big, you're paying 10x on >95% of your tokens.
For Pro, it's even worse, input/output is 4x and cache is 40x. The price different is really brutal. Yes you will still come out ahead by spending your first 10$/month on opencode go, but you will be saving a lot less than initially appears from their (60 USD for 10 USD pitch).
[0]: Deepseek: https://api-docs.deepseek.com/quick_start/pricing/ [1]: Opencode: https://opencode.ai/docs/zen/#pricing
Not true. Sol on XHigh or Max runs out even on the $200/mo plan. It's not close to effectively unlimited. Maybe at 2x the current allowance it can.
Real work. $200 looks good on the outside until the essence of it, e.g. the models lie. I gave a list of spec to Sol and Sol decided some items didn't need to be done and the reason was "unproven", "not enough evidence", etc.
They all come up with amazing ways to lie (or be lazy). Often times what you get isn't what you asked for (only on the surface). E.g. I ran it to iteratively bench and optimize a better data structure for the project. It spent hours and finally came up with something. When I check it out -- it benchmarked the wrong criteria and was way off. So here we go again. Most AI work looks good on the surface. There are infinite edge cases.
So to do real work and gate it you need to:
1. Plan
2. Get it to do the work
3. Get independent agents to check from different angles
4. Take that feedback and get it to fix those gaps
5. Match against the plan and redo parts if needed
Every task is easily 4-5x the estimated amount of tokens.
p.s. well I did burn some banked resets building a compiler for some language AND it is still NOT done. Every time it says done I say check it says ok we still have bugs...
> I'm running it in Oh My Pi with a second instance running as "advisor" and even with 5-6 active sessions (effectively 12 streams)
And to be frank, it is not that much weaker for regular software development work. I use Claude at work and I see no difference in capability. I only notice a dramatic difference in how much more expensive it is.