Seems to me we need tooling to do automatic offensive security as soon as a new frontier model comes out, with quickly turned around patches using the same frontier model. Rinse and repeat. Virtuous agentic security loop.
Every single time I've done this it has found at least one serious bug or security hole.
Makes me wonder how many are left I've not found.