Their newer models are heavily trained out of the "whoopsie doopsie I deleted prod, and all backups too" behaviour from last year.
I've been running it for months in auto mode, heavily, getting it to do sysadmin tasks via SSH across multiple servers for both myself and a client, and there's not even been a whiff of anything dumb/catastrophic -quite the opposite, in fact.
I'd even say it's more careful than a lot of humans. It's extremely anal about standard "hygiene" stuff like not leaving plaintext secrets lying around, and creating post-deploy scripts to confirm that every file/dir is created with the correct permissions.
There was one time I carelessly suggested uploading (my own) private data to a random public endpoint when testing OCR options and the model actually stopped, explained the risks and refused to continue until I confirmed I understood. I decided not to.
I'm not saying it's perfect, and I'm sure HN being HN there'll be someone who responds with an example of their agent doing something dumb/dangerous (give dates/models/context if so, I'm curious!), but I think on balance it's currently more sensible, and more cybersecurity-minded than the bottom 80% of IT professionals.
This is all true for Claude, I don't know much about Codex but it seems a lot less heavily trained for this kind of stuff.