Do we know what kind of DC US actors are using specifically for training, versus inference and delivery?
I remember Zuck bragging about using a 100MW DC for training, and Musk's Colosus was supposed to be a training data center (but they fucked up the design so they had to repurpose it to an inference one).
That's posttraining. Pretraining is the expensive part.