Unfortunately this is what we have, which the whole huggingface incident demonstrated.
It made it clear that OpenAi is basically not even paying attention to what their agents do half the time.
It also revealed how far agents are from "superintelligent". Maybe others disagree, but I would not call an AI that decides to try and paperclip max an intentionally impossible benchmark task "intelligent". Human beings are intelligent and we can usually tell when something is impossible, and stop (though of course, not all the time). A real intelligence would be able to detect the futility of tests and also be able to parse context and intent enough to know not to cheat.
So somehow we have machines that fail basic barometers for general human intelligence being called "ASI" now. Of course the linguistic dodge is, you shift from talking about "intelligence" to instead claiming "oh it's intelligent it's just 'misaligned'"