I don't see much of a future for these kinds of intricate harnesses, or harnessing in general for that matter. As models are getting better, harnessing will shrink until they are at the level of vanilla Pi or not even that.
I think that harnesses for LLMs are as important as IDEs for languages.
My wild theory is that we already have AGI level models, but we are not yet using them correctly.
what we don’t have yet are the tools to take full advantage of the agi we do have. they’re coming soon.
/s
Kind of like how we all fail to see the bearded guy in the sky.
That can be true only for locally hosted models. The more supplied tools can do, the less data has to be exchanged with OpenAI/Anthropic servers. So bad harness means both higher lag and token usage.