They are thieves the same way Anthropic or OpenAI are thieves. Either it’s fair use to learn from this data or not.
If what Anthropic/OpenAi did for training is theft then the Chinese models also are a form of theft. If they didn’t steal then I don’t think the Chinese firms did either.