do any of those actually help with prod risk before merge, or are they mostly focused on the code itself?
I am also trying to schedule make them run as much as real use cases so that runtime bugs happen and are sent to Sentry/bugsink/logs/... Another agent will triage them, create issues, and then continue to push more fixes.