Do local GPUs make sense? For the same price, can't you got a full years worth of cloud gpu time?
I would imagine that someone really serious about training (or any other CUDA workload) uses both.
If you only care about ML stuff, sure, the calculation is different.