So if you can get the spec right, and the LLM+agent harness is good enough, you can move much, much faster. It's not always true to the same degree, obviously.
Getting the spec right, and knowing what tasks to use it on -- that's the hard part that people are grappling with, in most contexts.
This is kind of how I feel. Chat as an interaction is mentally taxing for me.
At what cost,. monetary and environmental?
As costs drop exponentially (a reasonable expectation for LLMs, etc.) then increasing agent parallelism becomes more and more economically viable over time.
Not a reasonable expectation anymore. Moore's Law has been dead for more than a decade and we're getting close to physical limits.
But in all seriousness +1 can recommend this method.
Codex is like an external consultant. You give it specs and it quietly putters away and only stops when the feature is done.
Claude is built more like a pair programmer, it displays changes live, "talks" about what it's doing and what's working et.
It's really, REALLY hard to abort codex mid-run to correct it. With Claude it's a lot easier when you see it doing something stupid or getting of the rails. Just hit ESC and tell it where it went wrong (like use task build, don't build it manually or use markdownlint, don't spend 5 minutes editing the markdown line by line).
But I thought there are lots of agentic systems that loop back and ask for approval every few steps, or after every agent does its piece. Is that not the case?