Hard to imagine what a world with 100GW of compute looks like.
[1] https://epochai.substack.com/p/frontier-labs-dont-use-most-a...
^^ This quotes 1.4GW at the end of 2025. Add 0.3GW at Colossus 1, and some initial fraction of 1GW Trainium2 from [2]
Hard to imagine what a world with 100GW of compute looks like.
[1] https://epochai.substack.com/p/frontier-labs-dont-use-most-a...
^^ This quotes 1.4GW at the end of 2025. Add 0.3GW at Colossus 1, and some initial fraction of 1GW Trainium2 from [2]
I think token counts and GW are a gross over simplification here. Not all tokens are the same in the amount of GPU time they consume or the size of the GPUs they require or the amount of energy they consume. There's a huge optimization potential here once these companies get serious about consolidating the business they have and executing much more efficiently. Given enough time, these companies can heavily optimize their operations. Short term growth and not slamming the brakes on that is their primary concern.
I have been trying Claude Code with DeepSeek 4 apis, and the experience is barely different. In fact the margin of error is so small that harness and prompting account for the most impact in output quality.
But, here's the catch: I spend barely more than a handful of dollars per day of regular usage. In fact DS4 via api is cheaper than Claude 100$ subscription.
I really think that very soon many will start realizing that the alternatives are extremely close in performance but dramatically different in pricing.