Thanks for answering. That's really interesting.
Maybe agents flip this around. Instead of humans maintaining executable docs, the actual work generates the document.
It could become something like a PR-style review layer for agent work. You don't necessarily need to understand the underlying code or tooling, but you can inspect what changed, why it changed, and approve or reject it.
Do you think that would address any of the scaling problems you saw?