1,711 karma · joined August 16, 2009
I'm just happy there is capital behind the efficiency path because I am confident we can get positive value on that path.
It's a loop that uses adversarial review to check several dimensions of the writing:
https://github.com/Vibecodelicious/llm-conductor/blob/main/w...
The median household income in the US is $64000/year.
However at this point it can completely maintain itself. When a new version of Claude or OpenCode is released, it updates itself to work on the latest version. It can also add new implementations for harnesses pretty reliably. It's actually pretty fun to watch it at this point. "Make this work on Hermes agent and message me when you're done" and an hour later or so, I can go play with it in Hermes.
You also have some control over it. You can say something like "prune away the work we did on the UI. Make sure to remember that users can upload multiple images now"
In addition to these tools, I'm also building automation that will port the tools from the reference implementation (OpenCode) to other harnesses (Claude Code, Cline, Pi, Gemini, Kilo, Codex, others to come?). As well as automation that will either cherry-pick or re-implement commits onto the latest head from upstream.
[1]: https://github.com/Vibecodelicious/context-bonsai-agents#con...
[2]: https://blog.vibecodelicio.us/posts/how-i-fixed-context-wind...
https://www.youtube.com/live/z0JYVTAqeQM?si=oLvyLlZiFLTxL7p0
I hope it's something we can figure out how to propagate despite the seemingly limited interest. I suspect anyone who liked playing quake arena games would love this game if they are not susceptible to motion sickness.
I recently started exploring how to port open source shooters (red eclipse, warsow, nexuiz) to the platform and realized there are several considerations that make games designed for VR special that a pure port wouldn't hit.
I've also used tweakcc to make this work in Calude Code and plan to also do one for open source coding agents - codex, pi, Gemini, etc. And I'm also doing Livestreams of the development process.
I've honed a relatively decent flow that requires interaction from me for important parts (mostly) while making its own decisions at the not-important parts (mostly). This results in being able to send the agent off on an hours-long dev cycle and have relatively decent results after that need a few minor fixes. I think this is the best style for the current generation of AI
The feature allows the LLM to edit the context. For example, you can "compact" just portions of the conversation and replace it with a summary. Anthropic can see that the conversation suddenly doesn't share the same history as previous API calls.
In fact, I ported the feature to Claude Code using tweakcc, so it literally _is_ Claude Code. After a couple days they started blocking that with the same message that they send when they block third party tools.
[1]: the README.md describes the Context Bonsai features in my fork here: https://github.com/Vibecodelicious/opencode
As far as I understand, it's caches are not a "next-turn" thing, but a ttl thing.
I made the "retrieve" tool, which is what pulls back previously removed content, append to the conversation rather than putting it back where it previously was. But it's a but premature to really know if that's a real optimization.
I would really like people to try it out and report bugs, failures, and successes.
https://github.com/Vibecodelicious/opencode/blob/surgical_co...
I'm currently trying to get the LLM to be more proactive about removing content that is no longer useful in order to stay ahead of autocompaction and also just to keep the context window small and focused in general.