Won't the Chinese providers have to raise their prices as well due to the economics of serving inference at scale?
In reality, they're raising prices too. I was looking at Kimi K3 prices on OpenRouter and it's nearly the same as Anthropic and OpenAI.
Your other point is valid, but you're either vastly underestimating how many reads/writes a modern SSD can take or vastly overestimating how much output a typical LLM is capable of.
training cannot end for LLMs intrinsically. it's not some fixed cost. it's an ongoing one.
This is why LLMs are never going to be AGI. Humans don’t become obsolete just because they age.
(the point being that old, powerful professors have more than once blocked progress in their fields for decades. And only a funeral, eventually, solves the problem ...)