Sincere question: Has anyone figured out how we're going to code review the output of an agent fleet?
Models will improve, but also I predict code style and architecture will evolve towards something easier for machine review.
Initially we can absolutely just review them like any other PR, but at some point code review will be the bottleneck.