You didn't share yours and Adam's prompts in the post, so I'm left wondering how much of the success of this project is attributable to your collective ability and experience (both with this particular project and software in general) vs the capability of the model and harness itself? On that note, do you anticipate releasing LLMs at Oxide[1] (linked from RFD 0576)?
Personally I find credible success stories like yours interesting, if a little jarring. If they were commonplace, shouldn't software be generally getting a lot better?
[1] https://github.com/oxidecomputer/meta/tree/master/engineerin...