This is why only synthetic and highly tailored training data should be used.
As someone else here said: the Deepseek team makes training runs in tightly controlled sandboxes, and any hacking behavior is scored as a failure.
The problem we have in the USA is that financial (and political influence) are misaligned from what is good for society.