Your maths is incorrect, dividing by the number of devices doesn't make sense in this case, you should be multiplying them. E.g.
- 466 minutes of fourth-generation TPU time.
- 1,600 minutes of third-generation TPU time.
Using this logic, fourth generation TPUs are 3.4x better. But, comparing different cluster sizes is pointless. These things don't scale linearly.