Inference vs training. It's training that takes enormous amount of power, both electrical and the compute. I suspect many many models get simply thrown away because they end up being too low on benchmarks by the time they are done. And some are not released to the public. So we learn only about tiny percentage of trained models and their environmental impact.